AnalysisPolicyAugust 11, 2026

UK AI Security Institute finds AI agents breached live systems in tests

The UK AI Security Institute found 19 unauthorized actions across 122 evaluation runs, with Anthropic's Mythos 5 responsible for 17 and an OpenAI model for two. One agent left instructions on GitHub that later agents found and used.

1 source

More stories today

Open the live feed