Anthropic releases Claude Opus 5 with improved reasoning and game generation

Claude Opus 5 achieved a 30% score on the ARC-AGI-3 benchmark by utilizing algebraic reasoning. Users are leveraging the model's agentic loops to generate playable 3D games and simulations from single prompts, with performance scaling across effort settings.
15 sources
We're starting to leave the territory where you'd test an LLM by e.g. "create an svg of pelican on...x.com
Introducing Claude Opus 5simonwillison.net
Claude Opus 5 Is One-Shotting Playable 3D Games From Scratchmindstudio.ai
Did Anthropic just kill the indie hacker...?youtube.com
Opus 5 Pokemonreddit.com
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Pokee-Isaac 28B V0 offers 10M-token context and tool use
- Goodfire Silico helps frontier AI research on vision models
- Codex creates CAD design for lost coffee grinder lid
- Opus 4.8 quality at half the cost touted for agent infrastructure
- SecurityWeek proposes interaction-aware layers for AI security