Alibaba to release Qwen3.8-Flash-Next open-weight MoE model

Qwen3.8-Flash-Next (~125B-A6B + 51B n-gram) is an open-weight multimodal MoE model built on the architecture for the upcoming Qwen4 family. Ideal 4-bit quant ≈ 82 GB, with real-world quants likely 80–90 GB. Unsloth announced day-0 support.
6 sources
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Liquid AI open-sources Pipette benchmarking suite for on-device models
- wikiHow sues OpenAI over copyright infringement in AI training
- Claude Code 2.1.246 adds Auto mode tab, Bash wildcard warning
- Korean AI startup Wrtn raises funds at $870M valuation
- Podcast: Google DeepMind's Vivek Natarajan on AI in healthcare