CoreWeave says Vera Rubin NVL72 delivers 10x better tokens per megawatt
mark_k · x · 2026-07-23
CoreWeave says NVIDIA Vera Rubin NVL72 delivered about 10x more tokens per megawatt than GB200 Blackwell NVL72 on DeepSeek R1.
- The post cites CoreWeave’s first measured silicon results for Vera Rubin NVL72.
- It claims similar per-user responsiveness while cutting cost per token dramatically.
- The chart shows roughly 800,000 TPS/MW for Vera Rubin versus 80,000 TPS/MW for GB200 NVL72, framed as a 10x improvement.
- The implication is that this kind of efficiency jump could make large-scale, always-on reasoning and agent workloads more economical.
Related event: NVIDIA Launches Vera Rubin Platform with 10x Energy Efficiency(6 posts)→
More from Infra
- OpenAI reportedly lifts projected compute spending to about $750 billion by 2030 — Polymarket · 2026-07-23
- LLM apps need queues, caching, and sanity checks to survive async generation APIs — Defiant_Dentist5191 · 2026-07-23
- Compute is the Spice: Meta Eyes $10B Compute Lease Deal with Anthropic — thealexbanks · 2026-07-23
- Pinokio 8.0.40 hotfix checks Hugging Face token validity after login revocations — cocktailpeanut · 2026-07-23
- GLM-5.2 hits 12.2 tok/s on 16 AMD MI50s with llama.cpp RPC — Legal-Ad-3901 · 2026-07-23
- Developer Reflects on Local AI Hardware: Prepaid Workstations Reduce Experiment Friction — generativist · 2026-07-23