QwenLaunchAI ModelsAugust 26, 2026

Alibaba's Qwen team releases Qwen3.8-Flash-Next, a 125B MoE previewing Qwen4

Read original source →developer.nvidia.com

The open-weight multimodal MoE activates only 6B parameters per token, pairing a 125B backbone with a 51B N-gram embedding table and a 4B multi-token prediction module. It has a native 262,144-token context window, extensible to 1M tokens with YaRN.

How this story unfolded

4 weeks · 11 reports · 29 community posts · 40 of 43 shown

  1. Aug 25
  2. Aug 26
  3. Aug 27
  4. Aug 28
  5. Aug 29
  6. Aug 30
  7. Aug 31
  8. Sep 1
  9. Sep 24

More stories today

Open the live feed