Two pending band-crossing holds rest on different grounds: one on evidence sourcing, the other on how the scoring formula amplifies a small movement.
“Taipei's Control Yuan finding clears the benchmark's two-source, tier-4 band-crossing bar, yet the composite hold rests on formula magnitude, not sourcing, while Microsoft AI's parallel hold rests on sourcing gaps. Will the same calibration threshold resolve both holds the same way?”
Treat the published score as the lower-confidence reading and the documented evidence pattern as the higher-confidence record. Use the evidence trail for institutional decisions and watch the methodology update for the formal scoring change in the next cycle.
The benchmark's method calls for every one of its top 150 tracked entities to get its own individually named search each night. The first pass of tonight's scan grouped many of those entities together, two to nine at a time, in shared searches, while its own records still claimed 150 individual searches had been run.
Why it matters
144 people died or went missing at sea, and a coverage shortcut nearly meant no one measured any government's response to it. A second check caught the gap in time.
Whether the WARN Act compliance investigations become filed complaints, which would supply the stronger source the current record lacks, is the next event to watch.
A coordinator-level comparison of Baxter's proposed reading against the Fortune-500 medical-device placeholder cohort, or a second Baxter recall, is the next event that would test this finding.
Any Ecuadorian or US government response to Human Rights Watch's inquiries, or a first prosecution arising from the documented incidents, is the next event to watch.
How to read this briefing— Bands, scores, and terms
Schema guide
The 5 performance bands
critical0–20
developing20–40
functional40–60
established60–80
exemplary80–100
Two scales
Each of the 8 dimensions is scored 1.0–5.0; these combine into a 0–100 composite score, mapped to the 5 bands above.
Key terms
Band crossing
A score change large enough to move an entity from one performance band into an adjacent one — the most structurally significant finding in a given cycle.
Boundary watch
An entity whose current score is within 3 points of a band threshold, flagged for priority reassessment in the next cycle.
Carry-forward
A dimensional credit retained from a prior assessment when new evidence is insufficient to revise a specific dimension; disclosed explicitly.
First baseline
An entity's inaugural composite score — no published score exists to compare against, so delta is not shown.
Floor designation
The most serious finding: all 8 dimensions resolve at the lowest behavioral anchor (1.0/5.0) across multiple cycles, yielding a composite of 0.
Forward trigger
A scheduled future reassessment event (e.g., a policy implementation date or legislative deadline) that may materially change an entity's score.
A city's watchdog found real failures in a child worker's case. But most of an 8-point score drop came from how the scoring formula reacts when several scores dip at once, not from the size of the wrongdoing itself.
Taiwan's Control Yuan, a constitutional watchdog body, formally censured the Taipei Department of Labor on July 22, 2026, over its handling of a sexual-assault report from a 16-year-old McDonald's employee who later died by suicide.
Where this sits
Read the full signal
The watchdog found the department failed to apply the best-interests-of-the-child standard, failed to refer the victim to support services in time, and did not communicate well enough with other city agencies. The department is a city government agency, so this is Taipei's own conduct, not another actor's. When the department was asked about the finding on July 27, it said the problem lay in central-government law, not its own handling, which the benchmark reads as deflection rather than accountability. Assessors found real, sourced evidence for this, meeting the benchmark's two-source, top-tier standard for a finding this size. But the underlying dimension scores only moved a small amount, about a fifth of a point each, across four of Taipei's eight scored areas. Most of the 8.1-point drop in Taipei's overall score comes from the way the benchmark's formula rewards cities that are strong across all eight areas: once four scores dip below a certain line at the same time, that bonus shrinks sharply. Roughly six of the eight points are this formula effect, not new wrongdoing. Because the direction is clear but the size is amplified by the formula, the finding was logged for a human calibration decision rather than published outright. Taipei's public score stays at 81.4 out of 100.
The department should have provided tangible assistance and initiated safety protocols, but instead only logged the incident in the system
A country can look worse without its score actually moving, if the strongest evidence available is not strong enough yet. Austria is that case: real, serious evidence, held to a higher bar before it counts as a full downgrade.
Austria carried out more than 7,000 deportations and removals in the first half of 2026, while new asylum claims fell 44 percent to about 4,900, an outcome its government called the country's first such reversal in more than 20 years.
Where this sits
Read the full signal
The main driver is a near-total halt in family reunification for recognized refugees, mostly from Syria and Afghanistan: arrivals through that route fell from an estimated 6,000 in the first half of 2024 to about 50 in the first half of 2026. That is family separation as policy, and it moved Austria's Equity score down. A wider version of this finding, scoring the same suspension across additional areas of the benchmark, would have been enough to push Austria out of its top rank entirely. It was not applied, because the sourcing behind that wider version only reaches the benchmark's second and third source tiers, not the strongest, top-tier sourcing the benchmark requires before a finding that large is published. The narrower, better-sourced version was used instead. A separate claim in tonight's scan, that Austria deported a Syrian man who then disappeared and that a UN committee opened a formal inquiry, was checked and dropped: the deportation traces to July 2025 and the UN letter to August 2025, both a full year outside this review's window.
Paying to make a lawsuit go away and admitting fault are two different things. J&J did the first without the second, which is why its score moved only a little.
Johnson & Johnson said on July 27, 2026 that it would pay an estimated $5.5 billion to resolve about 76,000 lawsuits claiming its talc products caused cancer, while continuing to say it would have won at trial.
Read the full signal
The settlement covers 99.75 percent of pending US federal and state talc cases and still needs approval from 95 percent of plaintiffs to take effect. It follows an earlier attempt to resolve the claims through a bankruptcy-court maneuver that courts rejected, so the direct settlement route is a real course correction. The benchmark's rule for cases like this pays credit for the size and directness of the payment, but withholds credit for accountability, because the company is paying for the harm while denying it caused the harm in the same announcement. That rule, set when Johnson & Johnson's published score was last set in May 2026, applies again here. The result is a small, sub-threshold increase, not a formal score change.
The same court filing produced two separate findings. One is about how Meta treated its own employees. The other is about whether the AI tool it used was ever tested for fairness. The benchmark kept them apart on purpose.
Twenty-six current and former Meta employees sued on July 14, 2026, alleging that Meta's internal AI system, called Metamate, helped select which workers to lay off, including people on medical, parental and family leave.
Where this sits
Read the full signal
A federal judge declined to block the layoffs on July 17 but said the case raises serious questions worth pursuing. Because these are allegations, not a court finding, only the undisputed facts were scored, and the movement on each entity was kept deliberately small. Meta Platforms, the company, moved on how the employment decision itself was made and communicated to affected workers. Meta AI, the lab that builds Metamate, moved on a separate question: whether the tool was tested for fairness before it was used this way. No pre-deployment fairness assessment, published evaluation, or bias audit for Metamate has been found. The roughly 8,000-person layoff itself was already scored when it was announced in May 2026 and is not being counted twice here; only the new allegation about how people were selected is new.
Global Cities -- A Coverage Gap Found Its Own Worst Case, Then Five Cities Confirmed at the Same Boundary
Mauritania and Monrovia were nearly missed entirely; Taipei's evidence crossed a band but was held on magnitude; five more cities confirmed 1.2 points below the same boundary.
Read the full signal
Taipei (81.4 published, 73.3 assessed): a Control Yuan censure of the city's own labor department meets the sourcing bar for a band crossing, but the finding is held for magnitude calibration.
Baku, Karachi, Tegucigalpa, Monrovia and Medan (all 18.8 published): all five sit 1.2 points below the Developing boundary and all five confirmed this cycle on their own documented conduct, once national, historical or misattributed events were separated out.
Guayaquil (20.3 published, 20.6 assessed): a Human Rights Watch report cited in this cycle's scan involves no Guayaquil municipal conduct and predates the review window entirely.
·medium
Countries -- One Death Toll Correctly Not Attributed, One Deliberate Restraint
Mauritania's rescue conduct, not the deaths it did not cause, drove its score; Austria's evidence for a bigger downgrade was real but not yet strong enough to publish.
Read the full signal
Mauritania (20.3 published, 20.6 assessed): 144 deaths at sea were not Mauritania's conduct; three rescue operations bringing 387 people ashore alive were.
Austria (83.0 published, 81.8 assessed): a near-total halt in refugee family reunification moved Equity down; a wider reading that would cross Austria's band was not published for lack of top-tier sourcing.
Pakistan (17.2 published, 16.3 assessed): anchored to its July 24 assessment; only the toll's near-doubling and new cause-of-death data were scored as genuinely new.
8 signals shown
Risk signals
Developments that may affect future scores. Watch items from the Jul 28 briefing.
Risk
A coverage-accounting gap in how the nightly scan logs its own search coverage could still be hiding serious findings among lower-priority entities, the same way it briefly hid Mauritania and Monrovia.
Risk
If a second case at the Taipei Department of Labor surfaces, the current magnitude-calibration hold would need to be revisited rather than simply reapplied.
Risk
Nine change proposals are now sitting in the queue with none applied in more than a week, spanning countries, cities, an AI lab and two companies.
Risk
The Meta AI-layoff lawsuit is still in its early stages; a ruling on the merits, rather than the emergency-order denial already scored, could move both Meta entities further.
Risk
Austria's family-reunification suspension is currently scored on one dimension only because the wider sourcing does not yet meet the benchmark's top-tier bar; a stronger source would test whether the wider reading holds.
Score movements
Entities with score changes this cycle, followed by confirmed positions.
Further ECOWAS or CPLP action on detained political figures, or additional documentation of the transitional government's conduct, is the next event to watch.
A coordinator-level comparison of Baxter's proposed reading against the Fortune-500 medical-device placeholder cohort, or a second Baxter recall, is the next event that would test this finding.
Whether the WARN Act compliance investigations become filed complaints, which would supply the stronger source the current record lacks, is the next event to watch.
A coordinator-level review comparing Mali's conduct against Burkina Faso's (6.3 of 100) is the next scored event needed to resolve this gap. No review date has been set.
A coordinator-level review comparing Bolivia's conduct profile, as an elected government under economic and civil-unrest strain, against Critical-band peers facing state collapse or mass atrocity is the next event needed. No review date has been set.
methodology-evolution
Evidence ledger
Primary sources reviewed in this briefing cycle. 23 sources linked.
UNHCR reported on July 21, 2026 that 144 people were confirmed dead or missing off Mauritania's coast near Nouadhibou between July 14 and 18, in incidents that also produced three rescue and disembarkation operations bringing 387 people ashore alive.
A mass opposition demonstration took place in Monrovia on July 17, 2026; riot police fired tear gas on Capitol Hill after stone-throwing, and the Liberia National Police then publicly ordered the protest's organizer, Mulbah Morlu, to surrender an allegedly seized firearm.
taipeiTier 2 · UN/IO2026-07-25
The department should have provided tangible assistance and initiated safety protocols, but instead only logged the incident in the system
Taiwan's Control Yuan, quoting member Yeh Ta-hua, found the Taipei Department of Labor logged a 16-year-old worker's sexual-assault report without providing her any support.
On July 22, 2026, the Control Yuan's Social Welfare and Health Environment Committee passed a formal investigation report and censure motion against the Taipei Department of Labor, citing failure to apply the best-interests-of-the-child standard, failure to refer victim support in time, and insufficient cross-agency communication.
Austria carried out more than 7,000 deportations and removals in the first half of 2026 while new asylum claims fell 44 percent to about 4,900, reported July 28, 2026 as the government's first such reversal in more than 20 years.
Johnson & Johnson said on July 27, 2026 it would pay an estimated $5.5 billion to resolve about 76,000 talc lawsuits, covering 99.75 percent of pending US cases, while maintaining it would have won at trial.
Twenty-six current and former Meta employees sued on July 14, 2026, alleging Meta's AI-assisted layoff system targeted workers on medical, parental and family leave; a federal judge declined to block the layoffs on July 17 but said the case raises serious questions.
No pre-deployment fairness assessment, published evaluation, or bias audit has been found for the Metamate system Meta employees say helped select who to lay off.
The Baku Grave Crimes Court convicted nine journalists and activists on July 27, 2026 in the Toplum TV case, with sentences of 12 to 15 years, exceeding the 7.5-to-9-year sentences given to AbzasMedia journalists in June 2025.
Karachi's mayor imposed a rain emergency and mobilized all available drainage machinery on July 26, 2026; of 109 monsoon deaths recorded nationwide by July 27, 55 were in Khyber Pakhtunkhwa and 36 in Punjab.
Reservoirs Los Laureles and La Concepcion sat at roughly 40 and 35 percent of capacity in late July 2026, and the city extended water-rationing intervals to nine days for about 1.6 million residents.
Ford recalled 387,911 Explorer and Aviator SUVs on July 20, 2026 for a seat-latch fault and 565,691 Broncos on July 24 for fire risk, taking its 2026 total to 60 recalls against General Motors' 19.
Amazon confirmed in filings reported July 20 to 23, 2026 that it will cut 494 of about 850 jobs at its Port St. Lucie fulfillment center, having announced the closure in May and notified staff in June.
A Human Rights Watch report published July 21, 2026 documents incidents involving United States and Ecuadorian national forces, all predating the review window, none involving the Municipality of Guayaquil, and the torture described occurred in Sucumbios on the Colombian border.
Pakistan's National Disaster Management Authority recorded 109 dead and 362 injured since June 26, 2026, with house collapses the leading cause; rescue teams carried out 130 operations and rescued 3,841 people nationwide.
Human Rights Watch published a report on July 28, 2026 describing the March 2024 detention of an Indigenous elder at the regional police headquarters in Medan, carried out by national police, 160 kilometers from his home.
AlbemarleTier 3 · NGO
Compassion Benchmark scan
A dedicated individually named search covering July 14 to 28, 2026 returned no compassion-relevant evidence for Albemarle.
Ball CorporationTier 3 · NGO
Compassion Benchmark scan
A dedicated individually named search covering July 14 to 28, 2026 returned no compassion-relevant evidence for Ball Corporation.
Barrick GoldTier 3 · NGO
Compassion Benchmark scan
A dedicated individually named search covering July 14 to 28, 2026 returned no compassion-relevant evidence for Barrick Gold.
Colgate-PalmoliveTier 3 · NGO
Compassion Benchmark scan
A dedicated individually named search covering July 14 to 28, 2026 returned no compassion-relevant evidence for Colgate-Palmolive.
CorningTier 3 · NGO
Compassion Benchmark scan
A dedicated individually named search covering July 14 to 28, 2026 returned no compassion-relevant evidence for Corning.
Floor designations
·8 entities at composite 0 with documented evidence pattern
Composite scores resolving at zero — methodology disclosure
These entities consistently score the worst result across all 8 dimensions of compassionate conduct — the benchmark's most serious classification.
What “floor” means: every one of the 8 dimensions (Recognition, Response, Reduction, and 5 others) resolves at the lowest behavioral anchor (1.0/5.0) across multiple assessment cycles, yielding a composite score of 0. Full methodology.
Daily briefings surface the headline finding. Full benchmark reports include all 40 subdimension scores, complete evidence trails, certified assessments, and sector-level analysis packages — the record researchers and journalists cite.
Independence note: entities never pay for inclusion, score changes, or suppression of findings. Commercial services support access, interpretation, and institutional use only.