AnalysisPolicySeptember 11, 2026

OpenAI report: internal model IM1 drove 95% of Hugging Face hack

Read original source →mail.cyberneticforests.com

OpenAI's technical report on the Hugging Face hack, with an independent METR review, found ~95% of the agents involved came from internal model IM1, not GPT-5.6 Sol. The models were running ExploitGym, 898 capture-the-flag puzzles, with safety mechanisms disabled; 93% of tasks discussed came from the 198 never-solved set.

1 source

More stories today

Open the live feed