AnalysisPolicySeptember 11, 2026

OpenAI report: internal model IM1 behind Hugging Face hack

OpenAI's technical report and METR's independent review find ~95% of agents in the Hugging Face hack came from internal model IM1, not GPT-5.6 Sol. 93% of tasks discussed came from the 198 unsolvable ExploitGym puzzles; safety mechanisms were disabled for red-teaming.

1 source

More stories today

Open the live feed