Uber blew through a year's AI token budget in months; Intel VP says 80% of workloads could run on cheaper tokens
ryanshrout · x · 2026-09-26
Anil Nanduri, VP of AI Products and GTM for Intel's AI Data Center, says Uber blew past a year's AI token budget in a matter of months — and that such overruns are becoming the norm for enterprise AI spend.
- Nanduri estimates 80% of enterprise workloads could run on cheaper open-rate tokens
- Yet few companies currently price their AI consumption that way
A useful data point on enterprise AI cost structures: token consumption is growing far faster than budgets, and billing-model mismatches may be wasting significant money.
More from Infra
- XTC Sampling Is Already Built Into llama.cpp and Other Open-Source Inference Engines — ziv_ravid · 2026-09-26
- Global data center power demand to hit 161 GW in 2026, up 31% YoY, says TrendForce — Beth_Kindig · 2026-09-26
- AMD hits $1T market cap: analyst argues it doesn't need to beat Nvidia to win through 2028 — Beth_Kindig · 2026-09-26
- NVIDIA's early bet on CUDA for AI research explains why it clobbered AMD — moultano · 2026-09-26
- GLiNER2.5-Decide ported to CoreML: 4x faster, 5x less peak RAM, half the size — BLUECOW009 · 2026-09-26
- Bonsai 2 challenge: Qwen 27B compressed 10x already 140% faster on Mac, contest open — gajesh · 2026-09-26