Anthropic releases Claude Haiku 5.5, ~75% cheaper than Haiku 4.5
Read original source →anthropic.com
Anthropic's new small model costs around 75% less to run than Haiku 4.5 and is its fastest to date, built for high-volume tasks like summaries, compactions, and database queries. It is the first Haiku with an adjustable effort setting and pairs as a subagent with Opus 5.5 or Sonnet 5.5.
How this story unfolded
1 day · 9 reports · 17 community posts · 26 of 29 shown
- Oct 7
Claude Haiku 5.5 now available on AI Gatewayvercel.com
Anthropic launches Haiku 5.5 at a much lower pricethenewstack.io
Anthropic launches Claude Haiku 5.5 with 90% API price reduction, matching GPT-6 Lunaventurebeat.com
Introducing Claude Haiku 5.5 on AWSaws.amazon.com
Anthropic Launches Haiku 5.5: Its Cheapest and Fastest Claude Model Yetdecrypt.co
Claude Haiku 5.5anthropic.com
Claude Haiku 5.5 is now available in Devindevin.ai
- Oct 8
More stories today
Student builds Pamk desk robot with custom wake word
A French student built Pamk, a 3D-printed desk robot on an ESP32-S3 that manages tasks and reminders via the custom wake word "Hey Pamk," trained on thousands of the builder's own recordings. It needs no companion app.
r/robotics·1 hour ago
AT&T LegalEdge partners with OpenAI on legal AI workflows
AT&T's inhouse legal agency LegalEdge, launched in January, has handled nearly 100 matters in its first six months. OpenAI's Jason Boehmig said the effort centers lawyers' judgment, following OpenAI's Astra for Law launch and legal-tech plugin collection.
Artificial Lawyer·1 hour ago

Periodic Labs' Liam Fedus and Ekin Dogus Cubuk discuss AI scientists
Latent Space podcast interviews Periodic Labs' Liam Fedus and Ekin Dogus Cubuk on "synthesis superintelligence" — reinforcement learning grounded in physical experiments and autonomous labs. Periodic launched last September and is now building AI systems that reason over noisy physical experiments.
Latent Space·1 hour ago

Boston Dynamics' Spot automates inspections at Kirin Brewery Chitose Plant
Boston Dynamics·1 hour ago
Most AI SRE tools still leave incident decisions to humans
The New Stack webinar on October 22 features Traversal co-founder and CTO Raaz Dwivedi on why most AI SRE tools stop short of deciding when an incident needs attention. Dwivedi is also an assistant professor at Cornell Tech.
The New Stack·1 hour ago

AWS details sharing GPU clusters across teams with SageMaker HyperPod
AWS published guidance on partitioning expensive GPU clusters among multiple internal teams using Amazon SageMaker HyperPod, targeting generative AI workloads such as LLM training. The approach covers isolation boundaries, resource fairness, and operational independence for teams sharing the same hardware.
AWS AI Blog·1 hour ago

JetBrains releases Mellum2.1, a 12B MoE open model for coding agents
Mellum2.1 is a 12B mixture-of-experts thinking model activating 2.5B parameters per token, shipping under Apache 2.0 on Hugging Face with a 131,072-token context. It beats Mellum2 on 15 of 17 benchmarks and scores 82.0 on LiveCodeBench v6, but trails Qwen3.5-9B on Terminal-Bench 2.1 (17.4 vs 21.7).
MarkTechPost·1 hour ago

llama.cpp distributes inference across heterogeneous devices via ggml RPC backend
Georgi Gerganov·1 hour ago