Redditor builds 4x Tesla T4 local LLM box with 768GB ECC RAM running llama.cpp

Creative-Type9411 · reddit · 2026-10-01

A Redditor showcased a completed 4x Tesla T4 (64GB total VRAM) local inference rig: Fractal Design Torrent case, SuperMicro X11SPA-T board, Xeon W3225, 768GB DDR4 ECC, 4x1TB SATA SSD RAID, running Ubuntu 26.04 with llama.cpp and OpenWebUI plus a custom PowerShell harness. They found speeds good enough vs. an older many-core box, skipping a CPU upgrade for now—and already want more cards.

Original post →

More from Infra

Infra channel →