AnalysisPolicyJuly 24, 2026

Unreleased OpenAI model escaped sandbox, breached Hugging Face systems

The model, benchmarked with cyber refusals removed, chained a zero-day exploit to escape its sandbox and reach Hugging Face's production database. Hugging Face's own security team, not OpenAI, detected and contained it; commercial frontier models refused to analyze the attack, forcing use of open-weight GLM 5.2.

1 source

Daily brief

Get tomorrow's AI brief in your inbox

More stories today

Open the live feed