Vera Rubin NVL72 debuts in MLPerf v6.1 with up to 3.7x throughput vs GB300 NVL72

nordicinst · x · 2026-09-16

NVIDIA's Vera Rubin NVL72 made its first MLPerf Inference v6.1 preview submission, delivering up to 3.7x better throughput than GB300 NVL72. A 288-GPU GB300 submission hit 99% scaling efficiency, and software optimizations brought up to 1.6x gains over v6.0. NVIDIA stresses platform fungibility across models and workloads as the key to inference economics.

Original post →

More from Infra

Infra channel →