AnalysisAI ModelsSeptember 20, 2026

Reddit user benchmarks Qwen 3.8 27B on triple RTX 3090 setup

A r/LocalLLaMA user running Qwen 3.8 27B at UD-Q8_K_L across two RTX 3090s in tensor parallel reports it as the best fit for a 3x 3090 rig, after previously avoiding low-bit quantization and quantized KV cache.

1 source

More stories today

Open the live feed