Redditor crams six V100 GPUs into a standard full-tower case for local LLM inference
Odd_Caterpillar_2994 · reddit · 2026-09-18
A Reddit user shared a homelab build fitting six V100 16GB PCIe cards into a standard full-tower case instead of an open-frame chassis.
- Specs: EPYC 7262 CPU, ROMED8-2T motherboard, 6× V100 16GB PCIe
- Motivation: open-frame/server cases are large and unattractive
- Plan: currently testing, intends to run Qwen3 with TP2 + PP3 pipeline configuration for local inference
More from Infra
- First-ever PyTorch Day Japan lands in Tokyo on December 10, CFP open till Sept 27 — PyTorch · 2026-09-18
- Baseten launches server-side web search for open models, claiming 15% lower latency — AccBalanced · 2026-09-18
- Self-built Blackwell Colab pipeline generates a 2-hour MiniMax H3 movie for $6.96 — Interesting-Town-433 · 2026-09-18
- LaurieWired's CppCon keynote covers new memory hierarchies and how to prepare — lauriewired · 2026-09-18
- NVIDIA unpacks how CUDA's full stack powers specialized AI across finance, health and manufacturing — NVIDIA Developer · 2026-09-18
- 600 tok/s single-request on Qwen 35B with Ninfer on an RTX Pro 6000 — CharlesStross · 2026-09-18