Buying CMP mining GPUs for local LLMs? Bandwidth is the real bottleneck

StellarWox · reddit · 2026-09-01

A Redditor considering $1000 of NVIDIA CMP 100-210 mining cards for local LLM inference asks whether the cards' notoriously slow memory bandwidth matters if the model fully fits in VRAM, and whether multi-GPU layer-splitting only needs to pass activations between cards rather than weights.

Original post →

More from Infra

Infra channel →