OpenAI explores hiding model 'thinking', raising safety concerns

The Information reports OpenAI is testing a technique where models reveal less of their 'thinking', making them harder to monitor. Gary Marcus warns this could undermine chain-of-thought monitoring, a key safety tool.
1 source
Policy by email
Get an email when there's news on Policy
No news that day, no email.
More stories today
- Meta builds AI 'second brain' that learns from experts
- Snowflake shares surge 22% on AI coding momentum
- Gary Marcus critiques Musk's AI prediction shift
- Lutnick says Anthropic has patched relations with US government
- HPE lifts sales forecast on AI server demand