Huang's compute logic: Amdahl's Law caps speedups, over-provision the slow parts
le_james94 · x · 2026-09-29
From the Stanford CS153 thread series: Jensen Huang's design logic runs through Amdahl's Law — if 5% of a step can't be sped up, no number of GPUs gets you past 20x overall. So over-provision the slow parts. Midha proposed tokens-per-watt as the key metric, and Huang agreed.
Related event: Google: Frontier AI Labs Prefer Double Capacity Over Five Nines(3 posts)→
More from Infra
- Anthropic's IPO filing reveals potential $84.5B SpaceX compute bill — rohanpaul_ai · 2026-09-30
- Swift 1.5 matches Qwen3.8 27B quality on M5 Max while writing 34% fewer tokens — DerTomsn · 2026-09-30
- Anthropic inks up to $84.5B compute deal with SpaceX through 2029 — XFreeze · 2026-09-30
- general_compute partners with Cerebras to deploy 'world's fastest inference' — eptwts · 2026-09-30
- Virtual memory from first principles: page tables, TLBs, and Linux internals — abhi9u · 2026-09-30
- MIT's Daniela Rus: intelligence will live in your pocket, not a data center — MIT_CSAIL · 2026-09-30