LaunchAI ModelsJuly 15, 2026
Thinking Machines Lab releases Inkling, a multimodal MoE model

Inkling accepts text, image, and audio inputs and outputs text, with controllable inference effort balancing reasoning depth, token usage, and latency. It is available on Together AI's inference platform, served via an optimized FlashAttention-4-based kernel, and post-trained for scientific reasoning, coding, and agentic workflows.
15 sources
Together AI brings Thinking Machines Lab’s new model Inkling on day 0together.ai
RT @mervenoyann: Thinking Machines just dropped a ~1T Omni model 🔥 > 1M context window, trained...x.com
Thinking Machines just released with a ~1T param, 41B active, apache-2 model Benchmarks are a clear...bsky.app
Thinking Machines open sources first multimodal language model, Inkling, focused on low cost and 'resistance to censorship'venturebeat.com
What Is Inkling? Thinking Machines Labs' First Open-Weight Multimodal AI Modelmindstudio.ai
Mira Murati's New AI: Why Inkling Isn't Trying to Be #1youtube.com
Thinking Machines releases first open-weight model “Inkling”reddit.com
Thinking Machines by email
Get an email when Thinking Machines ships something
More stories today
- Thoughtworks' Kief Morris: humans must stay 'on the loop' in AI delivery
- GEMA wins major copyright ruling against Suno, orders damages paid
- LangChain builds ReviewBench benchmark for code review agents
- DeepSeek Flash 0731's reasoning trace amuses with 'OH MY GOD' outburst
- Former OpenAI VP Jerry Tworek discusses AI lab automation