QwenLaunchAI ModelsAugust 25, 2026

Alibaba releases Qwen3.8-Flash-Next, previewing Qwen4 architecture

Qwen3.8-Flash-Next is a multimodal MoE with 125B parameters plus 51B N-gram embeddings, activating only 6B per token. It beats Claude Opus 4.6 Max on 8 of 9 comparable benchmarks, with a 262K native context extensible to 1M. QwenCloud API pricing: $0.16/1M input, $0.47/1M output tokens.

How this story unfolded

5 days · 12 reports · 34 community posts · 46 of 49 shown

  1. Aug 25
  2. Aug 26
  3. Aug 27
  4. Aug 28
  5. Aug 29
  6. Aug 30

Qwen by email

Get an email when Qwen has news

No news that day, no email.

More stories today

Open the live feed