NVIDIA DSX squeezes 24% more token throughput from the same power budget, Lambda validates
NVIDIA Blog · rss · 2026-09-16
NVIDIA detailed its DSX AI factory platform at AI Infra Summit. Key results: Lambda ran 19 nodes within a 16-node power budget on HGX B200 servers, boosting cluster token throughput 24% (4M→5M tokens/s) and perf-per-watt 23%. NVIDIA's Eos factory, running Emerald AI's Conductor, has responded to 200+ demand-response signals from Silicon Valley Power, automatically cutting power from 4MW to 3MW while protecting priority jobs. DSX spans MaxLPS (dynamic power allocation), Flex (grid-responsive load balancing), OS, Sim, and reference designs with 800VDC. NVIDIA projects up to 40% more GPU capacity for Vera Rubin NVL72 factories within the same megawatt budget.
More from Infra
- Hitachi Energy to invest $528 million in new transformer factory in Mississippi — oilmutt · 2026-09-16
- Latham & Watkins, No.2 US Law Firm, Buys Nvidia Hardware to Fine-tune Open Weights In-house — MikeBirdTech · 2026-09-16
- Anthropic, Fluidstack and Cipher pledge $10M to fix a Texas town's water system — MxMnr · 2026-09-16
- Oracle CFO says she 'really, really' dislikes 'doing more with less' a day after layoffs — mkheck · 2026-09-16
- Astra optimizes its own inference on Rubin chips, doubling throughput in 72 hours — bookwormengr · 2026-09-16
- Audio8 open-sources on-device ASR/TTS models down to 0.1B, including iPhone offline transcription — FinanceYF5 · 2026-09-16