User reports running DeepSeek v4 Flash on RTX 4090 with DDR5
r/LocalLLaMA user reports running DeepSeek v4 Flash locally on an RTX 4090 with 128GB DDR5 5600 MT/s, using unsloth's UD-Q2_K_XL quant and latest llama.cpp on Ubuntu 26.04 with nvidia-595 driver.
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Vercel launches eve, its 'Next.js for agents' framework
- Black Hat 2026 wrap-up stresses open frameworks for agentic security
- Together AI adds educational documentation for LLM development concepts
- Auto mode launches after many months of internal use
- JPMorgan sees tech bond sales topping $500B as AI debt binge expands