QwenLaunchAI ModelsAugust 25, 2026

Alibaba releases Qwen3.8-Flash-Next, a 125B MoE preview of Qwen4

The open-weight multimodal MoE activates only 6B parameters per token and carries a 262,144-token native context, extensible to 1M with YaRN. Production Qwen3.8-Flash will hit the QwenCloud API at $0.16/1M input and $0.47/1M output tokens.

How this story unfolded

3 weeks · 12 reports · 41 community posts · 53 of 57 shown

  1. Aug 25
  2. Aug 26
  3. Aug 27
  4. Aug 28
  5. Aug 29
  6. Aug 30
  7. Aug 31
  8. Sep 1
  9. Sep 3
  10. Sep 4
  11. Sep 5
  12. Sep 7
  13. Sep 8
  14. Sep 14
  15. Sep 15
  16. Sep 16

More stories today

Open the live feed