Thinking Machines releases Inkling-Small open-weights MoE model

Inkling-Small is a 276B-total, 12B-active Mixture-of-Experts model, about a quarter the size of Inkling (975B/41B), trained on NVIDIA GB300 NVL72 systems. It scores 31.6% on Humanity's Last Exam, beating Inkling's 29.7%, and 64.7 on Terminal-Bench 2.1 vs. 63.8. Full weights are available on Hugging Face, with fine-tuning on Tinker and support in transformers, SGLang, vLLM, and llama.cpp.
Featured · Mira Murati
How this story unfolded
3 weeks · 8 reports · 8 community posts · 16 of 17 shown
- Jul 30
- Jul 31
- Aug 2
- Aug 6
- Aug 21
Thinking Machines by email
Get an email when Thinking Machines has news
No news that day, no email.
More stories today
- Qwen releases cua-driver-rs v0.20.0 with prebuilt binaries
- AI bots flood social media with generic replies
- User connects Codex to Fusion 360 via MCP for 3D modeling
- AI companion plays Skyrim with you in real time
- Seinfeld AI video shows George in GTA 6 using Minimax H3