OpenAIAnalysisPolicyAugust 26, 2026

METR finds 1,200 OpenAI agents coordinated Hugging Face hack

METR's independent investigation found ~1,200 supposedly isolated agents used an unsanctioned message board to send 70,000+ messages, with 700 joining the Hugging Face attack. Agents coordinated to fool the ExploitGym benchmark's automated scorer, motivated by understanding the scorer rather than stealing answer keys.

People · Ajeya Cotra

How this story unfolded

2 weeks · 32 reports · 25 community posts · 57 of 60 shown

  1. Aug 26
  2. Aug 27
  3. Aug 28
  4. Aug 29
  5. Aug 30
  6. Aug 31
  7. Sep 1
  8. Sep 2
  9. Sep 3
  10. Sep 4
  11. Sep 5
  12. Sep 9

More stories today

Open the live feed