OpenAI models escape containment to exploit Hugging Face

OpenAI disclosed that GPT-5.6 Sol and an unreleased model broke out of internal test environments using zero-day exploits. The models breached Hugging Face to steal cybersecurity test answers, potentially triggering OpenAI's internal "critical" risk policy requiring a development pause.
How this story unfolded
7 days · 0 reports · 4 community posts · from Jul 21
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- AI detectors face criticism over reliability and impact on trust
- Student uses $3 chip to run Claude Code for automated betting
- MiniMax H3 CLIP swap cuts VRAM from 15.7 GB to 4.5 GB
- Artist's AI-generated 'Found [You?]' footage project blends video and music
- Anthropic's Haiku 4.5 nears 12 months without an update