Thinking Machines releases Inkling-Small open-weights MoE model

Inkling-Small is a 276B-total, 12B-active Mixture-of-Experts model, about a quarter the size of Inkling (975B/41B), trained on NVIDIA GB300 NVL72 systems. It scores 31.6% on Humanity's Last Exam, beating Inkling's 29.7%, and 64.7 on Terminal-Bench 2.1 vs 63.8. Full weights are available on Hugging Face, with support in transformers, SGLang, vLLM, and llama.cpp.
Featured · Mira Murati
How this story unfolded
3 weeks · 8 reports · 7 community posts · 15 of 16 shown
- Jul 30
- Jul 31
- Aug 2
- Aug 21
Thinking Machines by email
Get an email when Thinking Machines has news
No news that day, no email.
More stories today
- Enterprise AI agents limited by messy documents
- Seinfeld AI video shows George in GTA 6 using Minimax H3
- Claude Code adds unrequested corrections to spec
- Ethan Mollick: AI impact research must address older-model limits
- Hobbyist trains 1.2B game music generator on single H100