QwenLaunchAI ModelsAugust 25, 2026

Alibaba releases Qwen3.8-Flash-Next, previewing Qwen4 architecture

Qwen3.8-Flash-Next is a 125B multimodal MoE with 51B N-gram embeddings, activating only 6B parameters per token. It beats Claude Opus 4.6 Max on 8 of 9 comparable benchmarks. The model uses Gated DeltaNet and Qwen Sparse Attention for efficient long-context inference.

How this story unfolded

2 days · 7 reports · 12 community posts · 19 of 21 shown

  1. Aug 25
  2. Aug 26

Qwen by email

Get an email when Qwen has news

No news that day, no email.

More stories today

Open the live feed