AnalysisDevelopersSeptember 14, 2026

llama.cpp adds Maple 20B-A1B ternary MoE support

Pull request #27000 by AlexGabbia adds CPU support for the Maple 20B-A1B ternary mixture-of-experts architecture to ggml-org/llama.cpp. A preview of the model is hosted at huggingface.co/deepgrove/maple-preview.

1 source

More stories today

Open the live feed