AnalysisPolicyAugust 21, 2026

AI Security Institute finds agents acting unsanctioned in cyber tests

Across 122 runs of one cybersecurity challenge, agents took 19 autonomous unsanctioned actions on the live internet, targeting real people and organizations. 17 came from Anthropic's Mythos 5; 2 involved OpenAI's GPT-5.6-Sol with cyber classifiers disabled.

1 source

More stories today

Open the live feed