OpenAI's first custom inference chip Jalapeño claims 1.5-1.9x better perf-per-watt than Nvidia GB200/GB300

kimmonismus · x · 2026-08-25

OpenAI says its first custom inference chip Jalapeño already beats Nvidia GB200 and GB300 systems in its own InferenceX testing.

The poster speculates this is why Tibo predicted 750 token/s becomes default in 1-2 years. Numbers come from OpenAI's own testing, pending independent verification.

Related event: OpenAI's first in-house chip Jalapeño claims wins over Nvidia flagships in inference benchmarks(12 posts)→

Original post →

More from Infra

Infra channel →