Anthropic releases Claude Sonnet 5.5, 30% faster than Sonnet 5
Read original source →anthropic.com
Sonnet 5.5 is the second model in the Claude 5.5 family after Opus 5.5, priced unchanged at $2/$10 per million tokens with $0.20 cache reads and a 1M context window. Anthropic says it costs up to 30% less for most work and is now the default Sonnet model on the API.
How this story unfolded
1 day · 16 reports · 54 community posts · 70 of 72 shown
- Sep 28
Introducing Claude Sonnet 5.5youtube.com
v2.1.284github.com
Anthropic releases Sonnet 5.5, which it calls a significantly cheaper, faster work partnertechcrunch.com
Anthropic launches Claude Sonnet 5.5 with 30% cost reduction per-task due to faster speeds and fewer tool callsventurebeat.com
Anthropic launches cheaper AI model, its second release since CEO's call for a slowdowncnbc.com
2.1.284code.claude.com
Claude Sonnet 5.5 now available on AI Gatewayvercel.com
Anthropic launches Claude Sonnet 5.5 with near-Opus performance at half the pricethenewstack.io
Introducing Claude Sonnet 5.5 on AWSaws.amazon.com
Anthropic's Claude Sonnet 5.5 Is Out, Beats Opus 5.5 at Coding for Half the Pricedecrypt.co
Claude Sonnet 5.5simonwillison.net
You picked Claude Sonnet 5.5 — but Anthropic may send your request to Sonnet 5thenewstack.io
- Sep 29
More stories today
NVIDIA Dynamo-Triton adds HSTU generative recommender inference
Dynamo-Triton now supports an end-to-end HSTU generative recommender workflow via NVIDIA's recsys-examples repo, combining PyTorch AOTI, FlexKV KV caching, and NV embedding cache. At batch size 8 on an RTX PRO 6000 Blackwell Workstation GPU, it hit up to 4.47x speedup for the three-layer HSTU model and 5.93x for the eight-layer model at 100% KV-cache hit rate.
NVIDIA Developer Blog·1 hour ago

METR's Chris Painter testifies to Senate on AI agent incidents
METR President Chris Painter testified September 30, 2026 before a Senate Homeland Security subcommittee hearing titled "Rogue AI: Securing the Homeland Against AI Agent Attacks." METR runs capability tests on frontier AI agents with voluntary access from OpenAI, Anthropic, Google, Meta, SpaceXAI, and Amazon, and says it is not paid or funded by them.
METR·1 hour ago

Google announces Gemini 4 Argon, rolling out first to cyber defenders
Gemini 4 Argon is Google's new frontier model, rolling out to trusted testers via the Fairwind Program before wider release. It scores 77.9% on DeepSWE v1.1 and 77.5% on AutomationBench-AA, and is priced at $2/$10 per million input/output tokens during introductory pricing.
Google DeepMind·2 hours ago

LangChain, Modal and Cogent Security host AI Heist challenge at SF Tech Week
LangChain·2 hours agoRFK Jr. claims AI will free Americans from "medical tyranny"
At a MAHA event with VP JD Vance, Health Secretary Robert F. Kennedy Jr. said AI is "better informed than any doctor in the country" and urged Americans to use it for second opinions. Ars Technica tested his claims: Gemini said masks do reduce respiratory disease spread.
Ars Technica·2 hours ago

Framework opens preorders for AMD Ryzen AI Max 400 desktop with 192GB
Framework's DIY Edition desktop with AMD Ryzen AI Max 400 Series and 192GB memory is now available for preorder. The 192GB unified memory configuration targets local LLM workloads.
r/LocalLLaMA·2 hours agoNVIDIA expands xio-sig with cuObject, ships SCADA Server SDK
NVIDIA added cuObject to xio-sig alongside cuFile, in partnership with Google Cloud and Microsoft, and made cuObject client and server libraries generally available. The new SCADA Server SDK lets storage providers build servers that answer GPU-initiated requests; IBM showed a prototype integrating SCADA with IBM Storage Scale.
NVIDIA Developer Blog·2 hours ago

Claude Code 2.1.286 ships 88 CLI changes
Release adds permission-prompt counts like "2 of 5" when requests stack, plus mouse support for "N more" rows in fullscreen lists. Sessions now retry once on the previous model of the same tier when the Anthropic API refuses the resolved model.
Claude Code Releases·2 hours ago