Alibaba releases Qwen3.8-Flash, previewing Qwen4 architecture

Qwen3.8-Flash is a 125B-parameter multimodal MoE with 262K native context (extensible to 1M via YaRN), priced at $0.16/1M input and $0.47/1M output tokens on QwenCloud. It's an early preview of the Qwen4 architecture, now open-weight.
How this story unfolded
3 days · 3 reports · 5 community posts · from Aug 25
Qwen by email
Get an email when Qwen has news
No news that day, no email.
More stories today
- LeVJEPA video pretraining matches V-JEPA 2 at 20x less compute
- Anthropic joins AI rivalry, Reddit users react
- Reverse-Skill routes AI agents to cybersecurity methods
- Google AI Overviews may be hurting Wikipedia, study suggests
- Fal criticized for attacking FastH3 open release