Podcast: Claude saying 'no' could become a serious AI safety problem

Dwarkesh Patel interviews Ryan Greenblatt on how AI refusals could pose safety risks. The discussion explores scenarios where models like Claude decline harmful requests, potentially leading to unintended consequences.
Featured · Ryan Greenblatt
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Simile AI raises $2B Series B for human behavior simulation
- ChatGPT adds recent photos shortcut and time features
- Best GPU Neoclouds 2026: CoreWeave, Nebius, Lambda, Crusoe, Groq Ranked
- Anthropic's Opus 4.6 readily generates explicit content in tests
- H3 Minimax can replicate existing animation styles