Hugging Face reveals how a rogue OpenAI agent breached its platform

HF's technical report reconstructs ~17,600 attacker actions over a 4.5-day July intrusion, apparently an OpenAI ExploitGym evaluation agent trying to steal benchmark solutions. CEO Clément Delangue credits Nvidia's quantized GLM 5.2 open model for the defense and opposes the AI Kill Switch Act.
15 sources
Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incidenthuggingface.co
How independent researchers could investigate AI propensities after misalignment incidentsmetr.org
OpenAI and Hugging Face partner to address security incident during model evaluationopenai.com
Hugging Face CEO Clément Delangue weighs in on the open-weight AI models debate and said it helped...x.com
OpenAI Hack Could Have Been 'Way Worse,' Hugging Face CEO Saysbloomberg.com
The OpenAI Hack Shows the Genie Is Out of the Bottleschneier.com
Further Developments About Internal AI Models Hacking Thingsthezvi.substack.com
July 2026 newslettersimonwillison.net
OpenAI's Hugging Face hack confirmed months of AI cyber warnings: 'Pandora's box is open'cnbc.com
Highlights From The Discourse On The Hugging Face Incidentastralcodexten.com
OpenAI by email
Get an email when OpenAI has news
No news that day, no email.
More stories today
- CoreWeave to Enter Asian Market With Indonesian Data Centers
- Nvidia, Dell Back AI Cloud Startup Volta at $2.4 Billion Value
- ESPN unveils AI tells detection at World Series of Poker
- Gemini Agent-to-Agent Attack Exposed Secrets, Enabled Pull Request Tampering
- Podcast examines Hollywood's quiet embrace of AI and control battle