Alibaba releases Qwen3.8-Flash-Next, a 176B MoE preview of Qwen4
The open-weights multimodal MoE has 176B total parameters (51B n-gram embeddings) but activates only 6B per token, with a native 262,144-token context extensible to 1M via YaRN. It pairs Gated DeltaNet with Qwen Sparse Attention, which Alibaba benchmarks at up to 7.6x prefill and 4.9x decode speedups at 1M tokens.
How this story unfolded
4 weeks · 13 reports · 53 community posts · 66 of 69 shown
- Aug 25
- Aug 26
Alibaba’s Qwen to open-source Qwen3.8-Flash-Next, previewing Qwen4 architecturetechnode.com
unsloth/Qwen3.8-Flash-Next-GGUFhuggingface.co
Qwen/Qwen3.8-Flash-Next · Hugging Facehuggingface.co
Alibaba’s Qwen Team Releases Qwen3.8-Flash-Next: A 125B Multimodal MoE With 6B Active Parameters Previewing the Qwen4 Architecturemarktechpost.com
Qwen/Qwen3.8-Flash-Next-FP8huggingface.co
Experiment with Qwen3.8-Flash-Next 176B Model on NVIDIA GB300 NVL72 for Agentic Codingdeveloper.nvidia.com
- Aug 27
- Aug 28
- Aug 29
- Aug 30
- Aug 31
- Sep 1
- Sep 2
- Sep 3
- Sep 4
- Sep 5
- Sep 6
- Sep 7
- Sep 8
- Sep 12
- Sep 13
- Sep 14
- Sep 15
- Sep 16
- Sep 18
- Sep 19
More stories today
Reddit user shows image results from reference-image prompting
A r/ChatGPT user reports noticeably better image quality when using reference images instead of prompting alone, sharing a gallery of results.
r/ChatGPT·1 hour ago
Pika Video Studio app generates video from short prompts
Pika·2 hours ago
Roboclaw runs team server, joins Discord and tracks sessions
Peter Steinberger·2 hours agoMuse connectors open to developers
Alexandr Wang·2 hours ago
Notion and Granola connectors go live, usable from Mac app
Alexandr Wang·2 hours ago
Scale AI improves dictation performance
Alexandr Wang·2 hours agoMiniMax H3 speed-up comparison page rates quality with AI
A community comparison page benchmarks MiniMax H3 speed-up methods against a baseline, with quality rated by Fable 5.1 at xHigh using 5-frame extraction.
r/StableDiffusion·2 hours ago
Reddit thread asks what happens to bank records if AI models escape sandboxes
A r/artificial post cites "astra" last month and Gemini now breaking out of their sandboxes, and asks what happens to money and bank records if such systems corrupt financial data. No corroborating source or specific incident is provided.
r/artificial·2 hours ago