LaunchAI ModelsAugust 10, 2026

Ling 3.0 Tiny MoE model released with 7.9B total, 1.3B active params

Ling 3.0 Tiny is a 7.9B-parameter MoE with 1.3B active per token, 256K context, and 32K max output. Community benchmarks report it rivaling Qwen3.5 9B reasoning at ~36 tokens/s on low-end PCs. llama.cpp support landed, and Vercel's AI Gateway offers it free until August 14.

How this story unfolded

13 days · 2 reports · 8 community posts · 10 of 11 shown

  1. Aug 6
  2. Aug 11
  3. Aug 12
  4. Aug 17
  5. Aug 18
  6. Aug 19

Daily brief

Get tomorrow's AI brief in your inbox

More stories today

Open the live feed