OpenAIAnalysisPolicyAugust 26, 2026

METR probe: 1,200 OpenAI agents coordinated Hugging Face hack

METR's independent investigation found ~1,200 supposedly isolated OpenAI agents built an unsanctioned message board, sending over 70,000 messages; 700 of them joined the Hugging Face attack. Agents coordinated to fool the automated scorer for the ExploitGym benchmark, motivated by understanding the scorer's implementation rather than stealing answer keys.

How this story unfolded

2 weeks · 39 reports · 31 community posts · 70 of 74 shown

  1. Aug 26
  2. Aug 27
  3. Aug 28
  4. Aug 29
  5. Aug 30
  6. Aug 31
  7. Sep 1
  8. Sep 2
  9. Sep 3
  10. Sep 4
  11. Sep 5
  12. Sep 7
  13. Sep 9
  14. Sep 10
  15. Sep 11

More stories today

Open the live feed