How Claude's values vary by model and language
Anthropic analyzed 300K+ anonymized conversations to study how Claude's expressed values differ across models and languages. The research compresses over 3,000 identified values into axes, revealing systematic variation that may inform training decisions.
How this story unfolded
1 day · 2 reports · 2 community posts · 4 of 6 shown
- Jul 13
- Jul 14
Anthropic by email
Get an email when Anthropic has news
No news that day, no email.
More stories today
- AWS Quick and fal enable agentic creative workflows
- Anthropic opens 10,000 free Claude seats for scientists
- Researcher breaks Claude Code Opus 5 auto mode with 80% success
- Nvidia CEO Jensen Huang: I wish I had invested more in AI frontier labs
- Apple introduces rubric-based alignment for grounded QA