AnalysisAI ModelsAugust 5, 2026

Community releases GGUF quants for pruned Qwen3-VL-32B 'heretic'

Community GGUF quants for the pruned 'heretic' Qwen3-VL-32B start at 6.7GB after removing unused text-encoder parts. They run via a fork of City96's GGUF loader.

1 source

Daily brief

Get tomorrow's AI brief in your inbox

More stories today

Open the live feed