LaunchAI ModelsAugust 2, 2026

inclusionAI releases Ling-3.0-flash, a 127.5B-parameter model

127.5B total params, 5.1B active, 512 experts with 8 active per token; MIT license with BF16 (~255GB) and official FP8 (~128GB) weights on Hugging Face. Reddit testers report ~80 tok/s decoding on a single DGX Spark.

How this story unfolded

3 weeks · 2 reports · 12 community posts · 14 of 15 shown

  1. Jul 23
  2. Jul 25
  3. Jul 27
  4. Jul 31
  5. Aug 4
  6. Aug 5
  7. Aug 8
  8. Aug 9
  9. Aug 11

Daily brief

Get tomorrow's AI brief in your inbox

More stories today

Open the live feed