OpenAI's Jalapeño Chip Beats Nvidia GB200 in Efficiency

Beth_Kindig · x · 2026-08-30

OpenAI claims its new Jalapeño chip delivers 1.5X to 1.9X higher tokens per watt and 1.7X to 3.6X lower latency compared to Nvidia's GB200 and GB300 on models like GPT-OSS 120B, DeepSeek R1, and Kimi K2.5 1T at peak throughput.

Original post →

More from Infra

Infra channel →