Reddit user benchmarks Qwen 3.8 27B on triple RTX 3090 setup

A r/LocalLLaMA user running Qwen 3.8 27B at UD-Q8_K_L across two RTX 3090s in tensor parallel reports it as the best fit for a 3x 3090 rig, after previously avoiding low-bit quantization and quantized KV cache.
1 source
More stories today
Researcher criticizes Wiley AI disclaimer on accuracy
Abeba Birhane·3 hours agoPirate Face mirrors Hugging Face models as torrents
Pirate Face mirrors open models from Hugging Face as checksum-verified magnet links held peer-to-peer, so weights survive any single host going down. Each file carries its official Hugging Face SHA-256, and existing pipelines can point at the same paths and API with zero code changes.
Hacker News·3 hours agoCNBC: 'Robot relations' departments may emerge as AI spreads in workplaces
CNBC reports corporations are widely deploying AI — chatbots, humanoids, and automated management systems — with major effects on worker pay and autonomy. The piece suggests a future 'robot relations' function inside companies.
CNBC Technology·4 hours ago

Reddit demo turns memes into videos with VoxCPM2 and H3
A r/StableDiffusion post shows a meme-to-video workflow pairing VoxCPM2 for audio with H3 for video, using a seed photo sourced from Wikipedia.
r/StableDiffusion·4 hours ago
YuE2 open-source music model runs locally, challenges Suno
Reddit user ran the 7GB YuE2 model on a 16GB GPU, generating 48kHz FLAC audio versus Suno's 44kHz. The September-released open-source generator supports cover versions and text-to-song with unlimited output.
r/SunoAI·4 hours ago
SPEED v2 speeds up MiniMax-H3 diffusion without retraining
A rewritten SPEED implementation for MiniMax-H3 runs early diffusion steps at lower resolution, then progressively returns to full resolution, cutting compute without retraining. V2 adds more samplers and reduces the jank of the original release.
r/StableDiffusion·4 hours ago
CMP 170HX overclock hits 1.89 TB/s, doubles Qwen 3.8 27B token speed
Memory bandwidth on the 40GB CMP 170HX rose from 1,386.2 GB/s to 1,890.1 GB/s, a 36.4% gain. Qwen 3.8 27B token generation jumped from 110 T/s to 202 T/s with no other config changes.
r/LocalLLaMA·4 hours agoLeCun: auto-regressive LLMs alone won't reach human-level AI
Yann LeCun·4 hours ago