QwenLaunchAI ModelsAugust 25, 2026

Alibaba releases Qwen3.8-Flash-Next, previewing Qwen4 architecture

Qwen3.8-Flash-Next is a multimodal MoE with 125B parameters plus 51B N-gram embeddings, activating only 6B per token. It has a 262K native context (extensible to 1M with YaRN) and beats Claude Opus 4.6 Max on 8 of 9 comparable benchmarks. QwenCloud API pricing: $0.16/1M input and $0.47/1M output tokens.

How this story unfolded

4 weeks · 14 reports · 52 community posts · 66 of 70 shown

  1. Aug 4
  2. Aug 25
  3. Aug 26
  4. Aug 27
  5. Aug 28
  6. Aug 29
  7. Aug 30
  8. Aug 31
  9. Sep 1
  10. Sep 2
  11. Sep 3

More stories today

Open the live feed