RX 7900 XTX beats R9700 by 30% on GPT-OSS-20B local inference tests

glenbeer · x · 2026-09-16

Enthusiast 1337hero revisited the RX 7900 XTX for local LLM inference: it ran GPT-OSS-20B about 30% faster than R9700s, though Qwen3 8x27B (dense) numbers looked off, while Gemma4-26B-A4B Q80 came in fast — useful reference points for AMD local deployment.

Original post →

More from Infra

Infra channel →