Pairing a decade-old RX 480 with RX 7900 GRE boosts llama.cpp MoE inference 36% over single GPU

tabletuser_blogspot · reddit · 2026-09-08

A Reddit user benchmarked dual-GPU llama.cpp inference (Vulkan build 10453) with an RX 7900 GRE 16GB plus an old RX 480 8GB.

Original post →

More from Infra

Infra channel →