AINine safetyof isthe abiggest complex,AI evolvinglabs fieldjust withgot nograded universalon standardsafety —and termsgovernance like— alignment,and red-teamingthe andresults riskare thresholdsembarrassing. meanAnthropic differentled thingsthe topack differentwith developers.a GradingC+, frontierthree labs onoutright governancefailed, withoutand accountingnobody forearned theanything technicalclose nuanceto behinda safeguards,passing evaluationsmark. andFrontier deploymentAI contextsbuyers oversimplifiescan't ajust deeplytake layeredmodel challenge.cards Aat glossaryface ofvalue competinganymore; definitionsthey isn'tneed ahard failureanswers ofon therisk industrythresholds, —access itto reflectsexternal howtesting genuinelyand hardincident thisreporting problembefore isdeploying these systems.
Nine of the biggest AI labs just got graded on safety andis governancea —complex, andevolving thefield resultswith areno embarrassing.universal Anthropicstandard led— theterms packlike withalignment, ared-teaming C+,and threerisk labsthresholds outrightmean failed,different andthings nobody earned anything close to adifferent passing markdevelopers. FrontierGrading AIfrontier buyerslabs can'ton justgovernance takewithout modelaccounting cardsfor atthe facetechnical valuenuance anymore;behind they need hard answers on risk thresholdssafeguards, externalevaluations testingand accessdeployment andcontexts incidentoversimplifies reportinga beforedeeply deployinglayered these systemschallenge.
There's a 33% chance that the U.S. government will maintain a public list of incidents related to AI safety on Jan. 1, 2030, according to the Metaculus prediction community.
© 2026 Improve the News Foundation.
All rights reserved.
Version 7.4.1