Dual R9700 local inference: how much does PCIe Gen4 actually cost vs Gen5?

IngwiePhoenix · reddit · 2026-09-30

A builder who waited over two years before buying two AMD Radeon AI Pro R9700 GPUs (64GB VRAM total) plans to run llama.cpp with ROCm behind llama-swap for dual-GPU inference and VRAM pooling.

The open question: since DDR5 RDIMMs are prohibitively expensive, would a cheaper DDR4 + PCIe Gen4 platform significantly hurt inference performance compared to Gen5? The post is a help-seeking thread, useful for anyone planning budget local inference rigs.

Original post →

More from Infra

Infra channel →