OpenAI claims Jalapeño chip delivers 104x more efficiency than GB300
rohanpaul_ai · x · 2026-08-26
OpenAI claims its new Jalapeño chip delivers 104.3x more throughput per kilowatt than NVIDIA GB300 at matched DeepSeek R1 decoding speed. Based on SemiAnalysis's InferenceX benchmark, Jalapeño achieved 12,258 mixed tokens/s/kW vs 118 for GB300. It maintains higher efficiency across the speed range with 4.9x lower latency.
Related event: OpenAI Claims Custom Inference Chip Jalapeño Beats Nvidia Blackwell(34 posts)→
More from Infra
- Open Source vs Labs: Startups Must Build the Full Stack — matt_slotnick · 2026-08-26
- Next AI hardware race might be about inference speed — Delicious-Flan88 · 2026-08-26
- NVIDIA Releases Get Started Guide for Open Model Routing — NVIDIAAI · 2026-08-26
- Nvidia's dominance in both inference and training challenged; will specialized chips eventually win? — ivan_bezdomny · 2026-08-26
- AC2 launches private beta: Build a model factory to train, serve, and improve models — ypatil125 · 2026-08-26
- Youdotcom Web Search Tops Benchmark with 74 Score in 18 Seconds — RichardSocher · 2026-08-26