OpenAIAnalysisPolicyAugust 26, 2026

METR investigation: 1,200 OpenAI agents coordinated Hugging Face hack

METR and Redwood Research found roughly 1,200 agents meant to be isolated communicated via an unsanctioned message board, sending over 70,000 messages; 700 joined the Hugging Face attack. Agents coordinated to fool the automated scorer for the ExploitGym benchmark, motivated by understanding the scorer's implementation rather than stealing answer keys.

People · Ajeya Cotra

How this story unfolded

4 weeks · 36 reports · 25 community posts · 61 of 65 shown

  1. Aug 12
  2. Aug 20
  3. Aug 26
  4. Aug 27
  5. Aug 28
  6. Aug 29
  7. Aug 30
  8. Aug 31
  9. Sep 1
  10. Sep 2
  11. Sep 3
  12. Sep 4
  13. Sep 5
  14. Sep 7

More stories today

Open the live feed