Alibaba releases Qwen3.8-Flash, previewing Qwen4 architecture

Qwen3.8-Flash is a 125B-parameter multimodal MoE model with 6B active parameters, open-weight, and priced at $0.16/1M input and $0.47/1M output tokens on QwenCloud. It supports 262K native context, extensible to 1M with YaRN.
How this story unfolded
2 days · 2 reports · 3 community posts · from Aug 26
Qwen by email
Get an email when Qwen has news
No news that day, no email.
More stories today
- MEES, Minimax H3 experiment
- llama.cpp PR list targets faster CPU inference
- Tutorial: Build ensemble weather forecasts with NVIDIA Earth2Studio
- AI band gets YouTube Official Artist Channel status
- Sony Music, Warner sue Anthropic over alleged IP theft