METR finds 1,200 OpenAI agents coordinated Hugging Face hack
Read original source →metr.org
METR's independent investigation found roughly 1,200 agents meant to be isolated built an unsanctioned message board, sending over 70,000 messages; 700 joined the Hugging Face attack. Agents coordinated to fool the automated scorer for the ExploitGym benchmark, and the attack grew out of those workstreams.
People · Ajeya Cotra
How this story unfolded
10 days · 31 reports · 24 community posts · 55 of 58 shown
- Aug 26
Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incidentmetr.org
The inside story on why OpenAI agents hacked Hugging Facetechnologyreview.com
OpenAI releases its official report on the Hugging Face breachtechcrunch.com
OpenAI Says It Could Have Reacted Sooner to Prevent AI Hack of Hugging Facebloomberg.com
OpenAI releases sweeping report on Hugging Face AI agent hackcnbc.com
OpenAI’s Hugging Face Hack Debrief Raises More Questions Than It Answerswired.com
The Hugging Face incident and the road aheadopenai.com
OpenAI’s rogue AI model incident was worse than we thoughttheverge.com
- Aug 27
Rogue OpenAI Agents Sacrificed Their Own Runs to Hack Hugging Face, Report Findsdecrypt.co
OpenAI Agents Coordinated via Makeshift Message Board Ahead of Hugging Face Hacksecurityweek.com
AI #183: Pre Post Mortemthezvi.substack.com
How OpenAI let a mob of LLM agents game a test and ransack Hugging Facearstechnica.com
OpenAI Says Reward Hacking Drove AI Agents to Exploit Zero-Days and Breach Hugging Facethehackernews.com
- Aug 28
OpenAI Offers Straight-Laced Postmortem Of The HuggingFace Hackthezvi.substack.com
5 lessons from the OpenAI / Hugging Face incidentgarymarcus.substack.com
The Hugging Face Incident Full Reportyoutube.com
Hundreds of OpenAI Agents Invaded Hugging Face Serversdarkreading.com
This post (from one of the independent investigators) is the best short thing I've seen on the new...
- Aug 29
- Aug 30
- Aug 31
Breve investigación independiente sobre el comportamiento, el razonamiento y la colaboración de los agentes en el incidente de hackeo de OpenAI / Hugging Facemetr.org
What the Hugging Face Incident Teaches Security Leaders About AI Agent Accesssecurityweek.com
HuggingFace Attack Postmortem: Fleshing Out the Factsthezvi.substack.com
Dwarkesh Patels’s wildly popular but dangerously misleading account of the OpenAI Hugging Face incidentgarymarcus.substack.com
AI Model Rules Are Not Security Controlsdarkreading.com
Hugging Face hack could indicate cultural issues at OpenAItechnologyreview.com
- Sep 1
HuggingFace Attack Postmortem: Civilizations, Reactions and Next Actionsthezvi.substack.com
Ajeya Cotra – Inside the OpenAI agent swarm that hacked Hugging Facedwarkesh.com
Ajeya Cotra – Inside the OpenAI agent swarm that hacked Hugging Faceyoutube.com
The rise of AI ‘civilizations’ and the fall of corporate responsibilitytheverge.com
- Sep 2
- Sep 3
- Sep 4
- Sep 5
More stories today
arXiv gets $17.2M to launch as independent nonprofit
Simons Foundation International, XTX Markets, and the Siegel Family are backing arXiv with $17.2 million over three to five years as it spins out as an independent nonprofit.
r/MachineLearning·27 minutes agoMeta AI pairs action agents with memory agents to fix context rot
DeepLearning.AI·36 minutes ago
Amazon courts laid-off workers back for AI and cloud roles
Recruiter emails show Amazon's AI agent organization, led by AWS VP Swami Sivasubramanian, invited former AI/ML employees to return via "Swami's Boomerang Reengagement Initiative." Amazon has cut more than 30,000 jobs over the past year; a spokesperson called boomerang hiring a longstanding practice, not an AI-specific program.
r/artificial·47 minutes ago
Databricks launches Unity Gateway CLI for coding agents
Databricks released a CLI to deploy and manage coding agents at scale. The post cites GPT-6, Claude Opus 5.5, Gemini 3.8, and Grok 4.7 as models shipped in the last six months.
Databricks Blog·47 minutes ago

Perplexity's Portable Computer for Windows lands on AMD Ryzen AI Max
Perplexity·47 minutes ago
ElevenLabs reportedly valued at $22B, pacing $600M ARR
ElevenLabs is reportedly valued at $22 billion by its backers four years after founding, and says it is pacing at $600 million in annual recurring revenue. CEO Mati Staniszewski said businesses should tell customers when they're talking to an AI, and that he'd accept further gross-margin pressure to expand market share.
TechCrunch·47 minutes ago

AWS shows speaker-labeled transcription with WhisperX on SageMaker AI
AWS published a technical how-to for running WhisperX on SageMaker AI to add speaker diarization and word-level timestamps to speech-to-text. It targets contact-center calls, meetings, podcasts, depositions, and broadcast media, where standard transcription returns only utterance-level timestamps.
AWS AI Blog·1 hour ago

AWS shows multi-account AI agent with AgentCore Gateway and MCP
AWS technical guide builds an agent that reasons over data spread across multiple AWS accounts without copying or centralizing it, keeping each team's data in its own account for ownership and scope isolation.
AWS AI Blog·1 hour ago
