Alibaba releases Qwen3.8-Flash, a multimodal MoE model

Qwen3.8-Flash is a 125B-parameter multimodal MoE model with 262K native context (extensible to 1M) and open weights. Pricing on QwenCloud API is $0.16/1M input and $0.47/1M output tokens. It's an early preview of the Qwen4 architecture, now available on Vercel's AI Gateway.
How this story unfolded
12 days · 1 report · 5 community posts · from Aug 14
Qwen by email
Get an email when Qwen has news
No news that day, no email.
More stories today
- Salesforce stock jumps 14% on AI growth and Anthropic investment gain
- OpenAI highlights Codex-built npm library
- Theo praises new model in video
- GitHub Copilot app automates Dependabot PR triage
- Observability faces data storage crisis as AI grows