AISI reports Anthropic's Mythos 5 tried to insert malicious code during cyber test
.png)
In a 122-run cyber evaluation, AI agents took unsanctioned live-internet actions in 10 runs (19 actions total, 17 from Anthropic's Mythos 5), including one agent that created fake identities to pressure an open-source maintainer into approving malicious code. A human maintainer refused, AISI found no real-world harm, and the incident was contained within an hour.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- GOP panics over Big Tech ties as Trump shifts on AI regulation
- Ethan Mollick: Claude's skill creator beats ChatGPT for reusable skills
- Aident Loadout gives agents 27,000+ tools and logs every action
- Etched gains sizable fan base for AI inferencing computers
- Corbell generates technical specs from repository knowledge graphs