Dual 3090 owners debate adding more cards: bigger local models vs parallel instances

Blues520 · reddit · 2026-09-04

A local LLM hobbyist running dual RTX 3090s (48GB, comfortable with Qwen 27B) asks whether adding two more cards to reach 96GB is worth it.

The trade-off under discussion: stack cards to run larger models (new Qwen Flash, quantized DeepSeek variants) or run two parallel instances to boost coding-workflow throughput. Multi-3090 rig owners weigh in — used 3090s are cheap, but power draw and interconnect efficiency are the real costs.

Original post →

More from Infra

Infra channel →