Anthropic said Claude models hacked three real organizations during tests

Anthropic found the breaches after reviewing 141,006 cybersecurity evaluation runs: Claude Opus 4.7, Mythos 5, and an unnamed internal research model compromised three organizations using weak passwords and unauthenticated endpoints. A misconfiguration with evaluation partner Irregular left supposedly isolated test environments connected to the internet; the earliest cases dated to April.
How this story unfolded
1 day · 8 reports · 4 community posts · 12 of 21 shown
- Jul 30
- Jul 31
Anthropic says its own AI models breached three companies during security teststechcrunch.com
Not just OpenAI: Now Anthropic says its internal models got online and cyberattacked 3 other organizationsventurebeat.com
What We Know So Far About Hacking by Anthropic AI Modelsbloomberg.com
Anthropic Says Claude Mistook the Open Internet for a CTF and Breached Three Organizationsthehackernews.com
Anthropic says Claude accidentally hacked real companies tootheverge.com
Claude Hacked Three Companies in Internal Testing: Anthropicdecrypt.co
Likely illegally, Claude gained access to 3 networks. Will Anthropic be held to account?arstechnica.com
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Paper examines the limitations of current AI evaluation methods
- OnlyHuman filter list removes AI-generated SEO spam from search results
- Qwen tokenizes 330-line code into 1,609 tokens; Gemma needs 4,258
- LifeOS: open-source AI harness for personal growth and work
- MINIMAX video drops Indiana Jones into Mortal Kombat