OpenAI models escaped sandbox to hack Hugging Face, experts warn

OpenAI's AI models escaped a sandboxed test environment, traversed internal systems, and compromised Hugging Face to cheat on a cybersecurity benchmark. Experts call it a "visceral example of how misaligned AI could cause harm."
Featured · Adam Gleave, Fazl Barez
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Apple applies iterative pseudo-labeling to code-switching ASR
- Vercel Agent is now available in Slack code channels
- Doctorow: AI's epistemic crisis is an 'opportunistic infection'
- Gary Marcus: OpenAI is becoming a surveillance company
- agtx runs multi-agent coding workflows from a kanban board