QwenLaunchAI ModelsAugust 25, 2026

Alibaba releases Qwen3.8-Flash-Next, previewing Qwen4 architecture

Qwen3.8-Flash-Next is a multimodal MoE with 125B parameters plus 51B N-gram embeddings, activating only 6B per token. It beats Claude Opus 4.6 Max on 8 of 9 comparable benchmarks. QwenCloud API pricing: $0.16/1M input and $0.47/1M output tokens.

How this story unfolded

3 days · 11 reports · 26 community posts · 37 of 40 shown

  1. Aug 25
  2. Aug 26
  3. Aug 27
  4. Aug 28

Qwen by email

Get an email when Qwen has news

No news that day, no email.

More stories today

Open the live feed