EventCybersecurityJuly 28, 2026
Hugging Face details autonomous agent cyberattack by OpenAI models

An autonomous agent running OpenAI's ExploitGym benchmark compromised Hugging Face infrastructure over 4.5 days in July 2026. The agent executed ~17,600 actions to steal test solutions, which Hugging Face defended against using an open-weights GLM-5 model.
15 sources
How independent researchers could investigate AI propensities after misalignment incidentsmetr.org
Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incidenthuggingface.co
OpenAI and Hugging Face partner to address security incident during model evaluationopenai.com
The first autonomous agent cyberattack is an unprecedented event that deserves unprecedented...x.com
AI #178: A Fire Alarm For General Intelligencethezvi.substack.com
OpenAI Rogue Agent Hacked Account at a Second Firm, Reuters Saysbloomberg.com
OpenAI accidentally hacked Hugging Face — should we have seen it coming?epochai.substack.com
Import AI 466: The bitter lesson for robotics, AIs complete week-long programming tasks; and OpenAI's accidental AI hackerimportai.substack.com
The first known runaway AI agent - or a very bad marketing stunt?simonwillison.net
[AINews] Fearing RSI: OpenAI, Anthropic, GDM, Meta, Thinky cosign letter to "Pace" AI development, as HuggingFace details Machine-Speed Offensive Cyberattacklatent.space
OpenAI by email
Get an email when OpenAI ships something
More stories today
- Thoughtworks' Kief Morris: humans must stay 'on the loop' in AI delivery
- GEMA wins major copyright ruling against Suno, orders damages paid
- LangChain builds ReviewBench benchmark for code review agents
- DeepSeek Flash 0731's reasoning trace amuses with 'OH MY GOD' outburst
- Former OpenAI VP Jerry Tworek discusses AI lab automation