768GB of VRAM for less than one RTX 6000: a 12x CMP170HX budget inference rig

segmond · reddit · 2026-09-18

Redditor segmond built a 768GB VRAM inference rig from twelve 64GB CMP170HX mining cards for less than the price of a single RTX 6000 Pro, linked via fiber to a second rig for RPC memory expansion. Running GLM, DeepSeek, Qwen, Kimi and MiniMax models through vLLM or llama.cpp, he claims performance no single RTX 6000 or M3 Mac Studio can match, and urges others to pounce on compute deals as demand stays high.

Original post →

More from Infra

Infra channel →