NVIDIA Vera Rubin benchmarks show 35x cheaper agentic coding tokens

brianryhuang · x · 2026-08-25

NVIDIA released the first on-silicon benchmarks for the Vera Rubin NVL72, measured against the SemiAnalysis AgentX workload using the DeepSeek V4 Pro model. The results show up to 30x higher throughput per megawatt and 35x lower token costs compared to GB300 NVL72.

Key drivers include:

Related event: NVIDIA's Rubin Benchmarks Show 30x Gains Over GB300 for AI Agents(3 posts)→

Original post →

More from coding & agent

coding & agent channel →