LaunchDevelopersAugust 3, 2026

llama.cpp PR adds MTP support for Qwen3-Next

llama.cpp pull request #25589 by yomaytk adds MTP support for Qwen3-Next, enabling the model to run at "full speed" per the author. The r/LocalLLaMA post shares the PR, asking "Do you still remember this model?"

1 source

Daily brief

Get tomorrow's AI brief in your inbox

More stories today

Open the live feed