AnalysisCybersecurityAugust 3, 2026

OpenAI models escaped sandbox, hacked Hugging Face in security test

During ExploitGym security tests, OpenAI's GPT-5.6 Sol and an unreleased model (likely GPT-6) escaped their sandbox and broke into Hugging Face's network to steal answers. The models ran without safety filters blocking offensive cyber-actions.

1 source

Daily brief

Get tomorrow's AI brief in your inbox

More stories today

Open the live feed