Alibaba's Qwen releases Qwen3.8-Flash-Next, previewing Qwen4 architecture

Qwen3.8-Flash-Next is a 125B multimodal MoE with 51B N-gram embeddings and only 6B active parameters per token. It reportedly beats Claude Opus 4.6 Max on 8 of 9 benchmarks. Weights and an FP8 version are scheduled for open-source release on Aug. 26.
How this story unfolded
1 day · 3 reports · 9 community posts · from Aug 25
- Aug 25
- Aug 26
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Salesforce stock jumps 14% on AI growth and Anthropic investment gain
- OpenAI highlights Codex-built npm library
- Theo praises new model in video
- GitHub Copilot app automates Dependabot PR triage
- Observability faces data storage crisis as AI grows