Building a DDR4+HBM2 local inference box: capacity over speed

Ancapta · x · 2026-09-03

A Reddit user asks which local LLM hardware upgrades actually feel meaningful, arguing that going from 32+16GB to 64+16GB is less worthwhile than 32+32GB. He plans to build a DDR4 + HBM2 inference machine to complement his main 32GB DDR5 + 16GB GDDR7 system, trading speed for greater total capacity to run a wider variety of models, and asks whether 32+32, 128+32 or 64+64 is the most logical capacity target.

Related event: Reddit Debates RAM vs VRAM Upgrades for Local LLMs(2 posts)→

Original post →

More from Infra

Infra channel →