Qwen3.8-27B open-weight model tops Hugging Face trending
Alibaba released Qwen3.8-27B under Apache 2.0, a 27B-parameter multimodal dense model with 262K native context (extendable to 1M via YaRN). It reportedly matches Opus 4.6 on some benchmarks and runs locally on an RTX 5090.
How this story unfolded
4 weeks · 2 reports · 46 community posts · from Aug 14
- Aug 14
- Aug 15
- Aug 16
- Aug 17
- Aug 18
- Aug 19
- Aug 20
- Aug 21
- Aug 23
- Aug 24
- Aug 30
- Aug 31
- Sep 1
- Sep 2
- Sep 3
- Sep 4
- Sep 5
- Sep 9
More stories today
OpenAI uses its models to fix critical vulnerabilities in own systems
Greg Brockman·33 minutes ago
Users debate letting Claude agents handle email
Reddit users discuss whether they trust Claude or other AI agents to triage real email inboxes, citing concerns about accuracy over time and the need to double-check filings. Few report letting agents fully manage email.
r/ClaudeAI·34 minutes ago
Terence Tao warns AI is depleting math's open problems
Terence Tao, a Fields Medalist, argues AI is solving fruitful open math problems faster than mathematicians can identify new ones. He suggests labeling problems "analysis-required" so bare AI answers without reasoning count for little.
Decrypt·53 minutes ago

Qwen by email
Get an email when Qwen has news
No news that day, no email.
NVIDIA Dynamo EPD disaggregation speeds multimodal serving up to 5x
NVIDIA's blog details EPD disaggregation for multimodal inference, claiming up to 5x faster time-to-first-token and 7x faster end-to-end response. It separates vision encoding from prefill/decode, best for image-heavy prompts and quantized MoE models.
NVIDIA Developer Blog·53 minutes ago

CUDA Toolkit 13.4 adds Windows on Arm support, Rubin preview
CUDA Toolkit 13.4 adds Windows on Arm support and early developer preview of the NVIDIA Rubin GPU architecture (compute capability 107). It also introduces MPS V3 with scriptable CLI, namespaces, and cgroup-integrated memory limits for shared GPU management.
NVIDIA Developer Blog·1 hour ago

Perplexity introduces Q2D-Web benchmark for agentic RAG retrieval
Perplexity AI·1 hour agoNew skills for interface review, typography, accessibility in Claude Code
Tom Doerr·1 hour ago
Claude tops new benchmark for agents that build agents, still fails most tests
On a new benchmark for AI agents that build other agents, Claude performed best but still passed fewer than a quarter of the tests. The benchmark, reported by The New Stack, highlights the difficulty of automating agent development.
The New Stack·1 hour ago
