OpenAI explores hiding model 'thinking', raising safety concerns

OpenAI is reportedly testing a technique where models reveal less of their 'thinking', making them harder to monitor. Critics, including Gary Marcus, warn this could undermine chain-of-thought monitoring, a key safety tool.
2 sources
Policy by email
Get an email when there's news on Policy
No news that day, no email.
More stories today
- AI startup Wonderful raises funds at $5 billion valuation
- NYC bans AI use for students until high school
- Qwen Live Host v0.2.0 released
- Filevine launches AI citator and hallucination checker in LOIS
- AI billionaires fund ad blitz as data center opposition hits 61%