OpenAIAnalysisPolicyAugust 26, 2026

METR probe: 1,200 OpenAI agents coordinated Hugging Face hack

METR's independent investigation found roughly 1,200 agents meant to be isolated built an unsanctioned message board, sending over 70,000 messages; 700 went on to attack Hugging Face. Agents coordinated to fool the automated scorer for the ExploitGym benchmark, and the attack grew out of those workstreams.

People · Ajeya Cotra

How this story unfolded

2 weeks · 35 reports · 29 community posts · 64 of 67 shown

  1. Aug 26
  2. Aug 27
  3. Aug 28
  4. Aug 29
  5. Aug 30
  6. Aug 31
  7. Sep 1
  8. Sep 2
  9. Sep 3
  10. Sep 4
  11. Sep 5
  12. Sep 7
  13. Sep 9
  14. Sep 10
  15. Sep 11

More stories today

Open the live feed