Alibaba releases Qwen3.8-Flash-Next, a 125B MoE preview of Qwen4
The open-weight multimodal MoE activates only 6B parameters per token and carries a 262,144-token native context, extensible to 1M with YaRN. Production Qwen3.8-Flash will hit the QwenCloud API at $0.16/1M input and $0.47/1M output tokens.
How this story unfolded
3 weeks · 12 reports · 41 community posts · 53 of 57 shown
- Aug 25
- Aug 26
Alibaba’s Qwen to open-source Qwen3.8-Flash-Next, previewing Qwen4 architecturetechnode.com
unsloth/Qwen3.8-Flash-Next-GGUFhuggingface.co
Qwen/Qwen3.8-Flash-Next · Hugging Facehuggingface.co
Alibaba’s Qwen Team Releases Qwen3.8-Flash-Next: A 125B Multimodal MoE With 6B Active Parameters Previewing the Qwen4 Architecturemarktechpost.com
Qwen/Qwen3.8-Flash-Next-FP8huggingface.co
Experiment with Qwen3.8-Flash-Next 176B Model on NVIDIA GB300 NVL72 for Agentic Codingdeveloper.nvidia.com
- Aug 27
- Aug 28
- Aug 29
- Aug 30
- Aug 31
- Sep 1
- Sep 3
- Sep 4
- Sep 5
- Sep 7
- Sep 8
- Sep 14
- Sep 15
- Sep 16
More stories today
Nvidia: AI security is an engineering problem across the agent stack
Nvidia argues AI security needs defined requirements, enforceable controls, named owners and evidence protections work, applied across models, harnesses and runtime environments. Example: an agent reading malicious instructions in an attached document should be blocked by network policy, with protected logs capturing the tool call and destination.
Nvidia AI Blog·1 hour ago

Unitree Dex5-S dexterous hand offers 22 degrees of freedom
Unitree's Dex5-S robotic hand packs 22 degrees of freedom and costs $6,500 per hand, per a Reddit post in r/Singularity. The Wuji Hand 2, a 20-DoF alternative, is priced between $16,000 and $20,000.
r/Singularity·1 hour ago
jevals replaces LLM judges with typed Jev decisions
Openlayer's jevals swaps free-form LLM-as-judge scoring for typed Jev decisions in evals. Posted as a Show HN project on GitHub.
Hacker News·1 hour agoHugging Face releases tokenizers v1
Hugging Face shipped tokenizers v1, a major version of its text tokenization library, with the release post focused on encode, decode, and scaling measurements.
Hugging Face Blog·1 hour ago
V7 uses GPT-5.6 to give AI agents institutional memory
V7 turns scattered company files into context agents can use for complex, source-linked work, built on GPT-5.6.
OpenAI Blog·1 hour ago

Engineer Travis adds AI voice to Big Mouth Billy Bass
Vaibhav Sisinty·1 hour ago
Databricks DevHub launches with agentic app templates
Databricks·1 hour ago
Healthcare AI needs "provenance of meaning" before learning from archives
Guest article by Heather Fricke argues archived clinical data can encode old access patterns, reimbursement incentives, and unequal treatment that AI will faithfully learn without understanding why. She cites NIST's healthcare AI work on data quality and calls for adding provenance of meaning to governance.
Healthcare IT Today·2 hours ago
