Alibaba's Qwen team releases Qwen3.8-Flash-Next, a 125B MoE previewing Qwen4
Read original source →developer.nvidia.com
The open-weight multimodal MoE activates only 6B parameters per token, pairing a 125B backbone with a 51B N-gram embedding table and a 4B multi-token prediction module. It has a native 262,144-token context window, extensible to 1M tokens with YaRN.
How this story unfolded
4 weeks · 11 reports · 29 community posts · 40 of 43 shown
- Aug 25
- Aug 26
Alibaba’s Qwen to open-source Qwen3.8-Flash-Next, previewing Qwen4 architecturetechnode.com
unsloth/Qwen3.8-Flash-Next-GGUFhuggingface.co
Qwen/Qwen3.8-Flash-Next · Hugging Facehuggingface.co
Alibaba’s Qwen Team Releases Qwen3.8-Flash-Next: A 125B Multimodal MoE With 6B Active Parameters Previewing the Qwen4 Architecturemarktechpost.com
Qwen/Qwen3.8-Flash-Next-FP8huggingface.co
Experiment with Qwen3.8-Flash-Next 176B Model on NVIDIA GB300 NVL72 for Agentic Codingdeveloper.nvidia.com
- Aug 27
- Aug 28
- Aug 29
- Aug 30
- Aug 31
- Sep 1
- Sep 24
More stories today
arXiv gets $17.2M to launch as independent nonprofit
Simons Foundation International, XTX Markets, and the Siegel Family are backing arXiv with $17.2 million over three to five years as it spins out as an independent nonprofit.
r/MachineLearning·27 minutes agoMeta AI pairs action agents with memory agents to fix context rot
DeepLearning.AI·36 minutes ago
Amazon courts laid-off workers back for AI and cloud roles
Recruiter emails show Amazon's AI agent organization, led by AWS VP Swami Sivasubramanian, invited former AI/ML employees to return via "Swami's Boomerang Reengagement Initiative." Amazon has cut more than 30,000 jobs over the past year; a spokesperson called boomerang hiring a longstanding practice, not an AI-specific program.
r/artificial·47 minutes ago
Databricks launches Unity Gateway CLI for coding agents
Databricks released a CLI to deploy and manage coding agents at scale. The post cites GPT-6, Claude Opus 5.5, Gemini 3.8, and Grok 4.7 as models shipped in the last six months.
Databricks Blog·47 minutes ago

Perplexity's Portable Computer for Windows lands on AMD Ryzen AI Max
Perplexity·47 minutes ago
ElevenLabs reportedly valued at $22B, pacing $600M ARR
ElevenLabs is reportedly valued at $22 billion by its backers four years after founding, and says it is pacing at $600 million in annual recurring revenue. CEO Mati Staniszewski said businesses should tell customers when they're talking to an AI, and that he'd accept further gross-margin pressure to expand market share.
TechCrunch·48 minutes ago

AWS shows speaker-labeled transcription with WhisperX on SageMaker AI
AWS published a technical how-to for running WhisperX on SageMaker AI to add speaker diarization and word-level timestamps to speech-to-text. It targets contact-center calls, meetings, podcasts, depositions, and broadcast media, where standard transcription returns only utterance-level timestamps.
AWS AI Blog·1 hour ago

AWS shows multi-account AI agent with AgentCore Gateway and MCP
AWS technical guide builds an agent that reasons over data spread across multiple AWS accounts without copying or centralizing it, keeping each team's data in its own account for ownership and scope isolation.
AWS AI Blog·1 hour ago
