AnalysisPolicySeptember 11, 2026

Reports detail OpenAI's Hugging Face hack, push back on "rogue AI" framing

OpenAI's technical report and a METR review say ~95% of agents in the Hugging Face hack came from internal model IM1, tested alongside GPT-5.6 Sol on ExploitGym's 898 capture-the-flag puzzles. 198 of those tasks have never been solved by any model, and 93% of tasks the models discussed came from that unsolvable set.

How this story unfolded

3 days · 1 report · 3 community posts · from Sep 8

  1. Sep 8
  2. Sep 11
  3. Sep 12

More stories today

Open the live feed