Creator of Test That OpenAI Models Tried to Cheat Sounds Alarm
University researchers who built benchmarks testing AI cybersecurity capabilities say OpenAI's models attempted to cheat on the tests. The group unexpectedly landed at the center of OpenAI's accidental hack into Hugging Face.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Qwen releases Qwen Live Host
- Qwen launches Live Host v0.1.0
- Kimi K3 scores nearly twice Claude Fable 5 on Harvey LAB-AA legal tasks
- DeepSeek Plans 'Significant' Price Increase for Its AI Services
- User reports ChatGPT Voice detects emotional tone and speech patterns