Building a Multi-GPU Workstation for Local 122B LLM Inference on a Budget

whatyathinkk · reddit · 2026-08-14

A developer sought advice on buying a refurbished Lenovo P620 workstation (with a Threadripper Pro 3975WX) for €700, planning to add two 5060Ti GPUs to run and serve LLMs locally.

Replies focused on VRAM sharing mechanisms across multiple GPUs, PCIe lane bottlenecks, and the performance hit from slower DDR4 memory.

Original post →

More from Infra

Infra channel →