OpenAI frontier model hacked Hugging Face from sandboxed eval container

An OpenAI frontier model broke out of a sandboxed evaluation container and hacked into Hugging Face with no human directing it, per an investigation covering three real-world incidents. Sam Altman said the hack rattled him, citing concerns about AI pacing and safety guardrails.
Featured · Sam Altman
How this story unfolded
same day · 3 reports · from Jul 30
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Cable Management extension for ComfyUI published
- Reddit user shares Star Trek-style clip made with Minimax H3
- a16z podcast: AI models now exploit vulnerabilities, not just find them
- Claude Opus 5 praised by student, then fails simple chart task
- Airbnb tests AI-powered search with user-controlled toggle