OpenAI's Jalapeño ASIC beats NVIDIA GB300 at half the power: 1.5-1.9x throughput per watt

新智元 · wechat · 2026-09-10

Anthropic has confirmed an internal custom silicon team for Claude inference chips, while OpenAI unveiled full benchmarks for its Jalapeño inference ASIC at HotChips: 700W vs NVIDIA's 1200W GB200 and 1400W GB300, with 1.5-1.9x higher throughput per watt and 1.7-3.6x lower end-to-end latency. The same week, Qualcomm signed a $60B custom inference chip deal with Amazon.

Original post →

More from Infra

Infra channel →