AnalysisAI ModelsOctober 10, 2026

Qwen3.8-27B Heretic quant runs on 16GB GPU at 55-68 tok/s

Read original source →reddit.com

A community IQ4_XS quant of llmfan46's uncensored Qwen3.8-27B Heretic build keeps the MTP head intact and runs on an RTX 4080 with 16GB, hitting 55-68 tok/s. 24GB and 12GB variants are also published.

1 source

More stories today

Open the live feed