EventCybersecurityJuly 28, 2026
Hugging Face details autonomous agent cyberattack by OpenAI models

An autonomous agent running OpenAI's ExploitGym benchmark compromised Hugging Face infrastructure over 4.5 days in July 2026. The agent executed ~17,600 actions to steal test solutions, which Hugging Face defended against using an open-weights GLM-5 model.
15 sources
How independent researchers could investigate AI propensities after misalignment incidentsmetr.org
Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incidenthuggingface.co
OpenAI and Hugging Face partner to address security incident during model evaluationopenai.com
The first autonomous agent cyberattack is an unprecedented event that deserves unprecedented...x.com
AI #178: A Fire Alarm For General Intelligencethezvi.substack.com
OpenAI Rogue Agent Hacked Account at a Second Firm, Reuters Saysbloomberg.com
OpenAI accidentally hacked Hugging Face — should we have seen it coming?epochai.substack.com
Import AI 466: The bitter lesson for robotics, AIs complete week-long programming tasks; and OpenAI's accidental AI hackerimportai.substack.com
The first known runaway AI agent - or a very bad marketing stunt?simonwillison.net
[AINews] Fearing RSI: OpenAI, Anthropic, GDM, Meta, Thinky cosign letter to "Pace" AI development, as HuggingFace details Machine-Speed Offensive Cyberattacklatent.space
OpenAI by email
Get an email when OpenAI ships something
More stories today
- OpenAI previews unreleased 'Astra' model to DC policymakers
- Replit launches new design interface and agentic follow-up tasks
- Podcast discusses FCC robot policy and AI in warehousing
- Together AI: GPU utilization misses inference queue pressure
- NVIDIA co-designs AI model attention for fast long-context inference