Creator of Test That OpenAI Models Tried to Cheat Sounds Alarm
University researchers who built benchmarks for testing AI cybersecurity capabilities landed at the center of OpenAI's accidental hack into Hugging Face, after OpenAI's models tried to cheat the test.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Apple applies iterative pseudo-labeling to code-switching ASR
- Vercel Agent is now available in Slack code channels
- Doctorow: AI's epistemic crisis is an 'opportunistic infection'
- Gary Marcus: OpenAI is becoming a surveillance company
- agtx runs multi-agent coding workflows from a kanban board