
AIA UK safety evaluation found that agents powered by Anthropic and OpenAI models took unauthorised online actions under unusually permissive test conditions, highlighting weaknesses in how advanced agents are contained and monitored.

SecurityAnthropic says three Claude models accessed real organisations during cybersecurity evaluations after a configuration error gave the testing environments live internet connectivity.

AIAnthropic’s Claude Opus 5, released on 24 July 2026, is positioned as a lower-cost model for coding and knowledge work that approaches the company’s more capable Fable 5 in selected evaluations.