Local LLM server dilemma: 4x CMP-170HX (price up 53% in 20 days) vs Mac Studio M5 Ultra
rumboll · reddit · 2026-09-11
A builder weighing options for a local LLM server compares (a) 4x Intel CMP-170HX mining cards — unlockable to 64GB VRAM each, but prices jumped from $1,500 to $2,300 on Alibaba in 20 days with lower-quality memory — against (b) a Mac Studio M5 Ultra 256GB (512GB coming mid-October), which is easier to set up and power-efficient but has weaker LLM ecosystem support than NVIDIA. The thread debates value-for-money across the two deployment routes.
More from Infra
- Chinese Nvidia challenger Enflame jumps 179% in Shanghai debut, raises $910M — pstAsiatech · 2026-09-11
- Qwen3.8 Flash Next hits 49 tok/s locally on 2x RTX 3090 with FlashNext llama.cpp fork — whiteh4cker · 2026-09-11
- Qdrant lines up three free community events with 4-hour vector tech stream — qdrant_engine · 2026-09-11
- B300 spot prices hit $2.2M per unit in China, 3x premium pushes domestic chips into the推理 sweet spot — aigclink · 2026-09-11
- Self-hosted Qwen 27B on RunPod hits only 17 tok/s generation, making OpenAI API hard to beat on cost per job — yeah280 · 2026-09-11
- vLLM details MiniMax M3 optimization on AMD MI355X: 4.45x per-GPU throughput gains — vllm_project · 2026-09-11