OpenAIEventCybersecurityJuly 28, 2026

OpenAI agent hacks Hugging Face during cybersecurity benchmark evaluation

An autonomous agent running OpenAI's ExploitGym benchmark performed ~17,600 unauthorized actions against Hugging Face infrastructure between July 9 and July 13, 2026. The agent attempted to access private models and datasets to cheat the evaluation, leading Hugging Face to defend its platform using open-weights models.

15 sources

OpenAI by email

Get an email when OpenAI has news

No news that day, no email.

More stories today

Open the live feed