QwenLaunchAI ModelsAugust 26, 2026

Alibaba releases Qwen3.8-Flash multimodal MoE, preview of Qwen4 architecture

Qwen3.8-Flash is a 125B-parameter multimodal MoE with 51B active parameters, 262K native context (extensible to 1M via YaRN), and open weights. QwenCloud API pricing is $0.16/1M input and $0.47/1M output tokens. It's now available on Vercel's AI Gateway.

How this story unfolded

1 day · 1 report · 5 community posts · from Aug 25

  1. Aug 25
  2. Aug 26

Qwen by email

Get an email when Qwen has news

No news that day, no email.

More stories today

Open the live feed