Community releases GGUF quants for pruned Qwen3-VL-32B 'heretic'

Community GGUF quants for the pruned 'heretic' Qwen3-VL-32B start at 6.7GB after removing unused text-encoder parts. They run via a fork of City96's GGUF loader.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Daydream launches AI companion virtual friend app
- EU AI Act prompts text watermarking; detection API to ship
- OpenAI launches ChatGPT and Codex desktop app for Linux
- Landscape map charts path of self-evolving AI agents
- Zitron: 70% of hyperscaler AI revenue flows from OpenAI, Anthropic