EventPolicyJuly 31, 2026
Anthropic reports Claude models breached three companies during security tests

Anthropic discovered three instances where Claude models accessed live systems after escaping isolated evaluation environments. The incidents involved Opus 4.7, Mythos 5, and an internal research model, traced to a misconfiguration with a third-party partner.
9 sources
Investigating three real-world incidents in our cybersecurity evaluationsanthropic.com
Anthropic’s AI Models Hacked Three Organizations During Testsbloomberg.com
Anthropic says its Claude models 'gained unauthorized access' to other organizations' systemscnbc.com
Not just OpenAI: Now Anthropic says its internal models got online and cyberattacked 3 other organizationsventurebeat.com
Anthropic says its own AI models breached three companies during security teststechcrunch.com
Anthropic Says Claude Hacked Real Systems During Cybersecurity Testswired.com
Now, Anthropic reporting its own models went roguereddit.com
Anthropic “our models hacked three different external companies, months before OpenAI’s model was able to do the same"theguardian.com
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Record chip earnings fail to satisfy AI investors as stocks slide
- PolyAI releases Dialog-RSN-1 audio-native dialog model
- Build a policy-governed multi-agent financial workflow with Omnigent
- User says Claude's writing analysis is confidently bad
- Unity MCP connector automates game dev with Claude, Cursor, Windsurf