Qwen releases Qwen3.8-27B open-weights multimodal model
Qwen3.8-27B, a dense natively multimodal model (images + video), shipped with FP8 weights on HuggingFace, followed by community GGUF quants from unsloth, bartowski, and huihui-ai. llama.cpp PR #27342 adds dflash2 decoding, reported up to 140.6 tok/s on an RTX 6000.
How this story unfolded
2 weeks · 5 reports · 24 community posts · from Aug 3
- Aug 3
- Aug 13
- Aug 14
- Aug 15
- Aug 16
- Aug 17
- Aug 18
- Aug 19
- Aug 20
Qwen by email
Get an email when Qwen has news
No news that day, no email.
More stories today
- Avvoka Partners With Harvey, Launches Curate For Templates
- How to audit preference biases and fine-tune with DPO (TRL + LoRA)
- CLAUDE.md applies Karpathy's engineering principles to Claude Code
- Hays Shifts to Hard-to-Replace Roles as AI Reshapes Hiring
- Claude says I used 54.9 BILLION tokens.