AnalysisAI ModelsSeptember 19, 2026

Qwen3.8-Flash-Next runs on 64GB Mac at ~27 tok/s

A 95.5 GiB Qwen3.8-Flash-Next checkpoint runs on a 64GB Mac by keeping routed experts on SSD and reading them only when a token routes to them. The author published the checkpoint and a fork after weeks of using it as a local coding model.

1 source

More stories today

Open the live feed