Nvidia claims Vera Rubin delivers 50X throughput per MW and 35X lower token cost vs Blackwell Ultra
Beth_Kindig · x · 2026-09-11
Analyst Beth Kindig cites Nvidia's claims that the next-gen Vera Rubin platform delivers 50X more throughput per MW and 35X lower token cost relative to Blackwell Ultra. She notes that if throughput gains outpace TCO growth, customers will be more willing to pay premium prices — the core of Nvidia's pricing power logic.
More from Infra
- 2.78T-param Kimi K3 runs inference on a single CPU in 8.24 GB of RAM — udmrzn · 2026-09-11
- Keep the Claude Desktop Workflow, Swap in Local Models via Ollama for Privacy — Technovangelist · 2026-09-11
- LithosAI ships Day-0 API inference for DeepSeek-V4.1-Flash at 250+ tokens/s per user — JiaZhihao · 2026-09-11
- TwelveLabs Marengo 3.0 Goes GA in Amazon Bedrock for Video Semantic Search — AWS ML Blog · 2026-09-11
- 1:26 continuous aerial AI video made entirely on a Mac with MiniMax H3 — cocktailpeanut · 2026-09-11
- KV cache gets QAT too: why this model beats others at fp4 KV cache — stochasticchasm · 2026-09-11