Qwen3.8-Flash-Next open-weights model previews Qwen4 architecture
Alibaba released Qwen3.8-Flash-Next, a multimodal MoE with 125B parameters plus 51B N-gram embeddings, activating only 6B per token. It has a 262,144-token native context, extensible to 1M with YaRN. QwenCloud API pricing: $0.16/1M input and $0.47/1M output tokens.
How this story unfolded
10 days · 14 reports · 54 community posts · 68 of 72 shown
- Aug 25
- Aug 26
Alibaba’s Qwen to open-source Qwen3.8-Flash-Next, previewing Qwen4 architecturetechnode.com
unsloth/Qwen3.8-Flash-Next-GGUFhuggingface.co
Qwen/Qwen3.8-Flash-Next · Hugging Facehuggingface.co
Alibaba’s Qwen Team Releases Qwen3.8-Flash-Next: A 125B Multimodal MoE With 6B Active Parameters Previewing the Qwen4 Architecturemarktechpost.com
Qwen/Qwen3.8-Flash-Next-FP8huggingface.co
Experiment with Qwen3.8-Flash-Next 176B Model on NVIDIA GB300 NVL72 for Agentic Codingdeveloper.nvidia.com
- Aug 27
- Aug 28
- Aug 29
- Aug 30
- Aug 31
- Sep 1
- Sep 2
- Sep 3
- Sep 4
More stories today
Insurers seek answers to rein in rogue AI
As incidents of unintended harm from rogue AI agents mount, CISOs and insurance firms are working out how to handle the fallout.
Dark Reading·1 hour ago

Ling 3.0 Flash Sante health model now free on Vercel AI Gateway
Ling 3.0 Flash Sante, a health-focused MoE model with 124B total parameters and 5.1B active per token, is free on AI Gateway through October 4. It features a 256K context window and function calling.
Vercel Blog·1 hour ago

Reddit user tests model with ostrich prompt
A Reddit user shared a 'fun little model test' on r/StableDiffusion, posting an image generated from the prompt 'Kshaturmurg (Ostrich)'. The post received 36 upvotes and 5 comments.
r/StableDiffusion·1 hour ago
Qwen by email
Get an email when Qwen has news
No news that day, no email.
AWS shows how to build a WhatsApp ordering assistant with Bedrock AgentCore
AWS blog post demonstrates deploying a multimodal WhatsApp ordering assistant using Amazon Bedrock AgentCore and Amazon Nova 2, targeting quick-service restaurants that manage ordering across multiple channels.
AWS AI Blog·1 hour ago

Zscaler CEO: AI drives more cybersecurity demand
Zscaler CEO Jay Chaudhry said AI is creating a major tailwind for cybersecurity, not reducing the need for software security. He noted CEOs, CIOs, and boards want AI for productivity but worry new models could create vulnerabilities.
Bloomberg Technology·2 hours ago

Perplexity details serving infrastructure for search models
Aravind Srinivas·2 hours ago
Perplexity details Ivy HTTP gateway for inference
Perplexity AI·2 hours agoReddit compares Sol 5.6 Ultra and Astra Light outputs
A Reddit user shared side-by-side outputs from Sol 5.6 Ultra and Astra 6 Light in work mode, calling both "insane" and saying they wouldn't need anything above Astra Light for real work.
r/Singularity·3 hours ago