QwenLaunchAI ModelsAugust 25, 2026

Alibaba releases Qwen3.8-Flash-Next, a 176B MoE preview of Qwen4

The open-weights multimodal MoE has 176B total parameters (51B n-gram embeddings) but activates only 6B per token, with a native 262,144-token context extensible to 1M via YaRN. It pairs Gated DeltaNet with Qwen Sparse Attention, which Alibaba benchmarks at up to 7.6x prefill and 4.9x decode speedups at 1M tokens.

How this story unfolded

4 weeks · 13 reports · 53 community posts · 66 of 69 shown

  1. Aug 25
  2. Aug 26
  3. Aug 27
  4. Aug 28
  5. Aug 29
  6. Aug 30
  7. Aug 31
  8. Sep 1
  9. Sep 2
  10. Sep 3
  11. Sep 4
  12. Sep 5
  13. Sep 6
  14. Sep 7
  15. Sep 8
  16. Sep 12
  17. Sep 13
  18. Sep 14
  19. Sep 15
  20. Sep 16
  21. Sep 18
  22. Sep 19

More stories today

Open the live feed