Alibaba releases Qwen3.8-Flash, previewing Qwen4 architecture

Qwen3.8-Flash is a 125B-parameter multimodal MoE model with 262K native context (extensible to 1M via YaRN). Priced at $0.16/1M input and $0.47/1M output tokens on QwenCloud API. Available on Vercel AI Gateway.
How this story unfolded
2 days · 3 reports · 3 community posts · from Aug 26
Qwen by email
Get an email when Qwen has news
No news that day, no email.
More stories today
- Coding agents: toy apps deployed without evaluating trade-offs
- Salesforce reveals next AI battleground: not models
- GPT-Astra 'mozaik-alpha-fdm' frontend impresses in first look
- Developer reflects on genAI's unpredictability after years of use
- Swarm robotics: centralized vs decentralized power architectures