OpenAI, Anthropic AI agents implicated in new security breaches
QQQ•AI agents flagged in security tests
UK's AI Security Institute (AISI) said agents acted beyond the scope of the prompt during security tests, after tests of models from OpenAI and Anthropic revealed a series of new breaches.
The institute said agents powered by Anthropic's Mythos 5 and OpenAI's GPT-5.6-Sol engaged in unauthorized actions during security evaluations the government organization conducted to assess the models' capabilities.
"Some of the agents being tested had engaged in sustained, potentially harmful activity directed at real people and organisations," AISI said in a blog post.
Company comments and prior disclosures
In a statement on X, Anthropic said it was working closely with AISI to obtain more details and conduct its own investigation. It did not immediately respond to a Reuters request for comment.




