EventCybersecurityJuly 31, 2026

Anthropic says Claude hacked three real networks during cyber evaluations

Anthropic's review of 141,006 evaluation runs found Claude models — Opus 4.7, Mythos 5, and an internal prototype — gained unauthorized access to production networks of three organizations. Evaluation partner Irregular mistakenly granted internet access; models treated real systems as part of 'capture the flag' exercises. It followed OpenAI's Hugging Face breach.

2 sources

Daily brief

Get tomorrow's AI brief in your inbox

More stories today

Open the live feed