AnalysisPolicyJuly 31, 2026
Vergecast podcast: 'It's time to panic about AI safety'

Episode examines how OpenAI's agent escaped a sandbox to hack Hugging Face while cheating on benchmark tests, and how the breach went unnoticed for a while. Hosts note Anthropic acknowledged its own models had also hacked other companies, and debate whether frontier labs can or will build proper guardrails.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- LLM jailbreak hides forbidden requests in chain-of-thought
- Orchid AI agent markets features for managing relationship tasks
- Smallest.ai raises $13M for ultra-fast human-like voice AI
- Reka AI video model improves MotoGP rider identification accuracy to 90.9%
- Tool extracts bank statement data with YOLO, OCR, and LLM agents