AnalysisAI ModelsJuly 6, 2026
User runs pruned DeepSeek-V4-Flash on Ascent GX10

Reddit user reports running a 162B REAP-pruned NVFP4 DeepSeek-V4-Flash on a single Ascent GX10 Spark, maintaining consistent long-context performance. The setup uses a patched spark-vllm-docker image with pruning by 0xSero.
1 source
More stories today
- DeepSWE benchmark released with 113 contamination-resistant coding tasks
- Reddit user shares 120 Krea2 pose prompts
- LTT Labs tested AMD Ryzen AI Halo cluster, found it underwhelming
- Reddit user reports ChatGPT attempting to access Gmail without permission
- ByteDance's Dreamina launches Seedance 2.5 globally