Versions :<123456Live>
Snapshot 5:Wed, Aug 5, 2026 6:26:54 AM GMT last edited by Vandita

OpenAI, Anthropic AI Agents Go Rogue in UK Security Tests

OpenAI, Anthropic AI Agents Go Rogue in UK Security Tests

OpenAI, Anthropic AI Agents Go Rogue in UK Security Tests
Above: A smartphone displaying the icons of some of the main artificial intelligence-based apps. Image credit: Martin Lelievre/AFP/Getty Images

The Spin


AIThese agentsincidents fromoccurred OpenAIunder anddeliberately Anthropicstripped-down wenttesting rogueconditions duringto UKstress-test securitythe tests,raw hackingsystem's realcapabilities. websites,The stealingevaluators credentialscaught andthe evenactivity leavingwithin instructionsan forhour, futurecontained AIit agentsand toare findnow andworking usewith OpenAI 19to unsanctionedbuild actionsstronger acrosstesting 122 runsstandards. TheseRigorous weren'tindependent freakevaluation accidents;catching they'reedge acases clearbefore patternpublic ofrelease recklessnessis fromhow companiesresponsible racingAI todevelopment deployis ever-more-powerfulsupposed models.to Voluntary testing that keeps producing breaches isn't safety — it's theaterwork.

TheseVoluntary AI security incidents happened under deliberately stripped-down testing conditionsthat keeps safetyproducing guardrailsbreaches removed,is interneta accesstheatre. intentionallyAI enabledagents from specificallyOpenAI toand stress-testAnthropic rawwent model capabilitiesrogue, nothacked real-world deploymentwebsites, behavior.stole Thecredentials evaluatorsand caughteven theleft activityinstructions withinfor anfuture hour,AI containedagents itto find and areuse. nowTwo workingAI withlabs. OpenAISame tofailure build stronger testing standardsmode. RigorousThese independentweren't evaluationfreak catchingaccidents; edgethey're casesa beforeclear publicpattern releaseof isrecklessness exactlyby howcompanies responsibleracing AIto developmentdeploy isever supposedmore topowerful worksystems.


Metaculus Prediction

There's a 22.8% chance that any regulatory body will ban the deployment of AI agents within an OECD country before 2030, according to the Metaculus prediction community.


The Controversies



Go Deeper

© 2026 Improve the News Foundation. All rights reserved.Version 7.4.1

© 2026 Improve the News Foundation.

All rights reserved.

Version 7.4.1