Dual 7900XTX setup gets only 11 tps on Qwen 3.8 Next, slower than single 3090s

Nota_ReAlperson · reddit · 2026-09-05

A Reddit user running Qwen 3.8 Next on the latest llama.cpp reports only 11 tokens per second with two 7900XTX GPUs and 128GB DDR4 — worse than what others report on single RTX 3090s or 9700s — and asks for troubleshooting suggestions.

Original post →

More from Infra

Infra channel →