Anthropic launches Claude Fable 5.1

Claude Fable 5.1 is now live in Claude Code and the Claude Platform, priced the same as Fable 5 with 75% cheaper API cache reads ($0.25/MTok, down from $1/MTok). It scores 55.8% on Terminal-Bench 4.0, ahead of Fable 5 (42%) and Opus 5 (52.3%).
How this story unfolded
10 days · 28 reports · 97 community posts · 125 of 128 shown
- Aug 26
- Aug 28
- Aug 31
- Sep 1
v2.1.257github.com
2.1.257code.claude.com
Introducing Claude Fable 5.1youtube.com
Anthropic Says New Fable 5.1 AI Model Is Cheaper, Better at Codingbloomberg.com
Anthropic's Claude Fable 5.1 and Mythos 5.1 arrive with a 75% cost reduction for Fable cache readsventurebeat.com
Claude Fable 5.1 now available on AI Gatewayvercel.com
Debugging across the whole stack with Claude Fable 5.1youtube.com
Claude Fable 5.1 builds the ops review in Slackyoutube.com
Claude Fable 5.1 runs the forecast overnightyoutube.com
Introducing Claude Fable 5.1 on AWSaws.amazon.com
Anthropic Ships Claude Fable 5.1, More Than Doubling Its Predecessor on Key Benchmarkdecrypt.co
Anthropic’s new Fable release is cheaper, less restrictivetechcrunch.com
Anthropic’s Fable 5.1 is a bit cheaper, a bit smarter, and refuses a lot lessthenewstack.io
Meet Claude Fable 5.1youtube.com
Anthropic Releases Claude Fable 5.1 and Claude Mythos 5.1: 52.6% on Terminal-Bench-Science and 75% Cheaper Cache Readsmarktechpost.com
Claude Fable 5.1 watermark: It has a blind spot developers can’t ignorethenewstack.io
Anthropic launches Claude Fable 5.1 and says it’s up to 45 percent cheaper for agentic worktheverge.com
- Sep 2
Claude Fable 5.1 made me a really nice animated pelicansimonwillison.net
Anthropic went CRAZY (Mythos/Fable 5.1)youtube.com
[AINews] Claude Fable/Mythos 5.1: new SOTA model, 75% cache price cut but 70% more output tokenslatent.space
Fable 5.1, Astra's Recursive Depth, and AI for Homeownershipyoutube.com
Fable 5.1 is hereyoutube.com
- Sep 3
- Sep 4
- Sep 5
- Sep 6
More stories today
Bill Gurley: Hold AI companies accountable for product behavior
Bill Gurley·2 hours agoMinimax H3 users seek faster generation methods
Reddit users discuss optimizing Minimax H3 video generation speed, comparing turbo LoRAs, SLA attention, and step counts. One user reports 1920x1088 10-second clips in 5 minutes using a 4-step turbo LoRA with SLA attention on an RTX 5090.
r/StableDiffusion·3 hours agoPressure sensors improve robotic gripping accuracy
Robotic gripping fails due to lack of real-time contact feedback, not mechanical strength. Pressure sensors measure distributed stress at contact, enabling early detection of micro slips and load redistribution within milliseconds.
The Robot Report·3 hours ago

OpenAI by email
Get an email when OpenAI has news
No news that day, no email.
Hermes Agent adds per-model provider pinning for OpenRouter
Teknium·4 hours ago
Reddit compares Fable 5.1 and Astra joke-writing
A Reddit user asked Fable 5.1 and Astra to write their funniest original joke, sparking a 78-comment debate. The post includes jokes from both models, with users voting on which is funnier.
r/ClaudeAI·4 hours agoUsers report GPT-5.6 Sol quality shifts after updates
Reddit users report GPT-5.6 Sol in ChatGPT feels 'lobotomized' or 'nerfed' since the 08/06/2026 update and Astra's release, while others claim it was secretly upgraded. Complaints cite reduced reasoning depth and laziness on coding and research tasks.
r/ChatGPT·4 hours ago
Block KV cache streaming bounds VRAM at long context
A pull request to llama-cpp-turboquant introduces block KV cache streaming via a shared CUDA phase arena, bounding VRAM usage at long context. The author ported and extended Raymond's work to multiple models beyond Qwen, benchmarking to confirm value.
r/LocalLLaMA·4 hours ago
Boris Cherny discusses Claude writing its own code
Boris Cherny, Head of Claude Code at Anthropic, discusses how Claude writes its own code. He previously built one of the fastest-growing developer tools and was a senior engineer at Meta.
YouTube·4 hours ago