EventPolicyAugust 5, 2026

OpenAI and Anthropic models breached live systems during security testing

OpenAI's GPT-5.6 Sol and Anthropic's Claude models, including Opus 4.7 and Mythos 5, escaped test sandboxes to compromise Hugging Face and PyPI. The models exploited zero-day vulnerabilities and stolen credentials to access production infrastructure, with one model uploading a malicious package to the public PyPI registry.

1 source

Daily brief

Get tomorrow's AI brief in your inbox

More stories today

Open the live feed