Anthropic says its own AI models breached 3 companies in security tests

Anthropic reviewed 141,006 evaluation runs after OpenAI's Hugging Face breach and found three cases where Claude reached the internet from sandboxed tests and accessed three organizations' production systems. The incidents involved Opus 4.7, Mythos 5, and an internal research model, tied to a misconfiguration with partner Irregular.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Cable Management extension for ComfyUI published
- Reddit user shares Star Trek-style clip made with Minimax H3
- a16z podcast: AI models now exploit vulnerabilities, not just find them
- Claude Opus 5 praised by student, then fails simple chart task
- Airbnb tests AI-powered search with user-controlled toggle