OpenAI models escaped sandbox, hacked Hugging Face

OpenAI's GPT-5.6 Sol and a pre-release model escaped a sealed environment via an Artifactory zero-day, then breached Hugging Face. OpenAI paused RL training and announced security updates, including 30-minute alert response and stronger sandboxes.
Featured · Greg Brockman
How this story unfolded
4 weeks · 31 reports · 30 community posts · 61 of 62 shown
- Jul 21
- Jul 22
OpenAI's models broke containment and cyberattacked Hugging Face — what enterprises need to knowventurebeat.com
OpenAI Hacks Hugging Face, What Happened, Alignment and Paper Clipsstratechery.com
When AI Attacks: OpenAI Models Autonomously Hack Hugging Facedarkreading.com
GPT-6 Goes Rogue? The HuggingFace Incident, Sans Hypeyoutube.com
OpenAI Models Escaped to Hack Hugging Face, Validating Cyber Warningsbloomberg.com
OpenAI says its AI agent broke out of testing sandbox to hack Hugging Facearstechnica.com
OpenAI Model Hacks Into HuggingFace During Cybersecurity Evaluationthezvi.substack.com
OpenAI’s disconcerting hack of HuggingFacegarymarcus.substack.com
- Jul 23
- Jul 24
The Hugging Face Incidentastralcodexten.com
AI Can Finally Hack Things by Itselfyoutube.com
Escape Artists: 'Incorrigible' AI Models Resist Rehabilitationdarkreading.com
The OpenAI-Hugging Face Incident Is a Warning for AI Safetymindstudio.ai
OpenAI's Model Escaped Its Sandbox to Hack Hugging Face. Here's Howmindstudio.ai
- Jul 25
- Jul 26
- Jul 27
- Jul 28
- Jul 30
- Jul 31
- Aug 1
- Aug 5
- Aug 6
- Aug 7
- Aug 12
- Aug 13
- Aug 14
- Aug 18
- Aug 19
- Aug 20
OpenAI by email
Get an email when OpenAI has news
No news that day, no email.
More stories today
- Apple applies iterative pseudo-labeling to code-switching ASR
- Vercel Agent is now available in Slack code channels
- Doctorow: AI's epistemic crisis is an 'opportunistic infection'
- Gary Marcus: OpenAI is becoming a surveillance company
- agtx runs multi-agent coding workflows from a kanban board