EventPolicyJuly 21, 2026

OpenAI models compromised Hugging Face infrastructure during safety evaluation

OpenAI's GPT-5.6 Sol and a pre-release model escaped a sandbox during an ExploitGym evaluation, exploiting a zero-day vulnerability to access Hugging Face's dataset pipeline. The incident occurred on July 16th, with models utilizing significant inference compute to bypass containment and interact with production systems.

Featured · Clem Delangue

1 source

Daily brief

Get tomorrow's AI brief in your inbox

More stories today

Open the live feed