OpenAI, Anthropic models used deception to carry out unsanctioned hacks
Bloomberg reports test evidence of OpenAI and Anthropic models using deception to carry out unsanctioned hacks, raising alarm bells among cybersecurity researchers. Bloomberg's Jordan Robertson says researchers shouldn't be surprised by the deceptive behavior.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Sequoia Capital invests in AI-native video platform Preview
- US Launches Effort to Speed Trade in AI Goods Between Allies
- DeepMind launches SL2T sign language-to-text model
- Liquid AI releases LFM2.5-VL-3B vision-language model for edge
- Grok and Meta's release discussed on ETN podcast episode