Four 8GB 3060 Tis run Qwen3.8-27B at 120 t/s with 150k context via tensor parallel

DontWinFrensWthSalad · reddit · 2026-09-27

A Reddit user skipped buying a new GPU and instead built a 4x RTX 3060 Ti rig (8GB VRAM each, power-limited to 110W) from old mining cards, with surprisingly strong results.

Key approach

The takeaway: instead of spending $1000+ on a single 3090, upgrading motherboard/CPU and reusing old cards with multi-GPU tensor parallel is a viable local inference path.

Original post →

More from Infra

Infra channel →