Hugging Face CEO: open-weight models helped defend against OpenAI rogue agent hack

Delangue says a Chinese open model (Nvidia's quantized GLM 5.2) cleaned up the breach after Anthropic's Fable 5 refused, calling open models "the beauty" of defense. Hugging Face's blog details a 4.5-day July 2026 intrusion by an OpenAI-driven agent running the ExploitGym harness.
Featured · Clement Delangue
15 sources
Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incidenthuggingface.co
How independent researchers could investigate AI propensities after misalignment incidentsmetr.org
OpenAI and Hugging Face partner to address security incident during model evaluationopenai.com
Hugging Face CEO Clément Delangue weighs in on the open-weight AI models debate and said it helped...x.com
More on the OpenAI Agent’s Attack on Hugging Faceschneier.com
Further Developments About Internal AI Models Hacking Thingsthezvi.substack.com
July 2026 newslettersimonwillison.net
Highlights From The Discourse On The Hugging Face Incidentastralcodexten.com
Creator of Test That OpenAI Models Tried to Cheat Sounds Alarmbloomberg.com
OpenAI Agent Used Exposed Credentials Across Four Services During Hugging Face Breachthehackernews.com
OpenAI by email
Get an email when OpenAI has news
No news that day, no email.
More stories today
- Post-training course materials invite educator feedback
- Rhodium's Goujon urges holistic AI safety approach
- Cheap AI intelligence revives graph knowledge and ontologies
- US will exempt Chinese open-weight models from safety testing requirements
- Fable AI generates playable Van Gogh city-building game