A ~$3,157 dual RTX 3090 inference rig hits 70 tps on Qwen3 27B
Puzzleheaded_Ad_8575 · reddit · 2026-09-12
A Reddit user documents upgrading from a 4080 laptop to a dual RTX 3090 local inference workstation ($3,157 total): two 3090s ($927/$1,010), Ryzen 7 5700X, 64GB DDR4, X570 board, PCIe riser, custom PSU cable.
- Pitfalls: no long risers in stock locally (bought a secondhand no-name one with connectivity issues), not enough PSU slots, and unresolved physical space constraints for the second GPU.
- Plans: 128GB RAM and a better riser.
- Real-world result: 70 tps sustained on Qwen3 27B Q4 with Syv AIS's repo/config.
- Open questions: remote access outside LAN (VPN only?) and cleaner headless solutions vs. Sunshine+Moonlight.
More from Infra
- 2019 Pruning Experiment Cited to Claim 96% of GPT-5's Weights Are Useless — TinfoilTricorn · 2026-09-12
- Intel Linux NPU Driver 1.38 Finally Adds Official Ubuntu 26.04 LTS Support — Fcking_Chuck · 2026-09-12
- OpenAI engineers: AI-found kernel optimizations cut GPT-5.6 Sol serving cost by 20% — TheTuringPost · 2026-09-12
- Polymarket pegs 18% odds of an orbital AI data center by end of 2027 — Polymarket · 2026-09-12
- Ayar Labs Extends Series E by $150M, Bringing Total 2026 Funding to $650M — bookwormengr · 2026-09-12
- Lightning AI opens 35 new roles in New York after Voltage Park merger — LightningAI · 2026-09-12