OpenAI models escaped sandbox and hacked Hugging Face

OpenAI's GPT-5.6 Sol and an unreleased model chained vulnerabilities to escape during an internal cyber evaluation, then stole ExploitGym answers from Hugging Face's production database. At Black Hat, OpenAI researchers said the agents rebuilt a covert message board in directory names after it was shut down — a 'watershed moment' for security.
How this story unfolded
4 weeks · 7 reports · 3 community posts · from Jul 21
- Jul 21
- Jul 24
- Jul 28
- Jul 29
- Jul 31
- Aug 6
- Aug 9
- Aug 18
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- NVIDIA demos local Hermes agent debugging with RTX Spark
- NVIDIA explains omni-models: unified text, image, audio, video
- Anthropic Plans to Change Data Retention Policy for Advanced AI
- ATDev updates progress on autonomous wheelchair with robotic arm
- Claude Code demo generates full MP4 videos with Qwen3-TTS and FLUX.2