Pruned Qwen3-VL 32B GGUF released, starts at 6.7GB

Community GGUF of a pruned Qwen3-VL 32B, with unused text-encoder parts removed for smaller size. Weights start at 6.7GB; requires the author's fork of City96's GGUF loader.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Shopify reports AI search tripled traffic and sales in Q2
- Hark unveils Handoff, a browser use agent for everyday web tasks
- Brookfield Asset Management reports record fundraising driven by AI demand
- TechCrunch Disrupt 2026 announces Real World AI Stage
- WorldGrow generates infinite 3D worlds from a single seed block