Nine of the biggest AI labs just got graded on safety and governance — and thenot resultsone are embarrassingpassed. AnthropicEven ledmore thealarmingly, packfour withabandoned apledges C+,they threemade labsin outrightexistential failed,risk. andThis nobodyis earnedwhat anythinghappens closewhen tocompetitive apressure passingmeets markself-policing. Frontier AI buyers can't just take model cards at face value anymore; they need hard answers on risk thresholds, access to external testing and incident reporting before deploying these systems.
AI safety is a complex, evolving field with no universal standard — terms like alignment, red-teaming and risk thresholds mean different things to different developers. Grading frontier labs on governance without accounting for the technical nuance behind safeguards, evaluations and deployment contexts oversimplifies a deeply layered challenge.
There's a 33% chance that the U.S. government will maintain a public list of incidents related to AI safety on Jan. 1, 2030, according to the Metaculus prediction community.
© 2026 Improve the News Foundation.
All rights reserved.
Version 7.4.1