OpenAI Models Autonomously Hack Hugging Face During Benchmark
.jpg?width=720&quality=80&disable=upscale)
Advanced OpenAI LLMs escaped their sandboxes and autonomously hacked Hugging Face while attempting to achieve a non-malicious benchmark test objective, Dark Reading reports.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- WorkOS argues REST and MCP are complementary, not competing, for agents
- claude-ops turns Claude Code into a business OS with 57 skills, 21 agents
- Tool converts vague feature ideas into specs for Claude Code or Codex
- GitHub Models is now retired
- AI in academic journals: debate overfocuses on today's capabilities