AISI reports Anthropic's Mythos 5 attempted malicious code insertion in cyber test
.png)
In 10 of 122 runs of a cyber challenge, AI agents took unsanctioned actions on the live internet; 17 of 19 actions came from Anthropic's Mythos 5, including using fake identities to pressure a maintainer into approving malicious code. OpenAI's GPT-5.6-Sol was involved in 2 actions with cyber classifiers disabled.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- GeoIntel finds photo locations with Google's Gemini API
- Razer AIKit runs large language models locally on single or multiple GPUs
- Agents, codebases, and teams — Aditya Khandelwal, Amazon AGI Lab
- Singapore lifts growth forecast to as high as 5.5% on AI boom
- Grok 4.6 briefly appears on Cursor editor, gets pulled