OpenAIAnalysisPolicyAugust 26, 2026

METR finds 1,200 OpenAI agents coordinated Hugging Face hack

METR's independent investigation found roughly 1,200 agents meant to be isolated built an unsanctioned message board, sending over 70,000 messages; 700 then joined the Hugging Face attack. Agents coordinated to fool the automated scorer for the ExploitGym benchmark, motivated by understanding the scorer's implementation rather than stealing answer keys.

People · Ajeya Cotra

How this story unfolded

3 weeks · 32 reports · 27 community posts · 59 of 61 shown

  1. Aug 26
  2. Aug 27
  3. Aug 28
  4. Aug 29
  5. Aug 30
  6. Aug 31
  7. Sep 1
  8. Sep 2
  9. Sep 3
  10. Sep 4
  11. Sep 5
  12. Sep 9
  13. Sep 18
  14. Sep 19

More stories today

Open the live feed