768GB of VRAM for less than one RTX 6000: a 12x CMP170HX budget inference rig
segmond · reddit · 2026-09-18
Redditor segmond built a 768GB VRAM inference rig from twelve 64GB CMP170HX mining cards for less than the price of a single RTX 6000 Pro, linked via fiber to a second rig for RPC memory expansion. Running GLM, DeepSeek, Qwen, Kimi and MiniMax models through vLLM or llama.cpp, he claims performance no single RTX 6000 or M3 Mac Studio can match, and urges others to pounce on compute deals as demand stays high.
More from Infra
- 28 hashes logged: crowd audits Gensyn's open-1b, first verifiable training run — benfielding · 2026-09-19
- Astonishing wealth transfer: hyperscalers' free cashflow halved, mostly to Nvidia — docmilanfar · 2026-09-19
- Token Superposition Training Cuts Pretraining Compute 2.5x on 10B MoE — gordic_aleksa · 2026-09-19
- Groq founder flags slowing US power production growth as AI buildout risk vs China — TheZachMueller · 2026-09-19
- Université Laval builds new AI-dedicated HPC cluster with Mila and Vector Institute — Mila_Quebec · 2026-09-19
- Virginia, world's largest data center hub, unveils new restrictions on large projects — Polymarket · 2026-09-19