FULL STORY
OpenAI's Jalapeño Chip Benchmarks Beat Nvidia Flagships
SemiAnalysis reveals OpenAI's in-house inference chip Jalapeño, with benchmarks claimed to beat Nvidia's GB200/GB300. OpenAI plans deployment by year-end as the first step of a multi-generation chip roadmap, positioning itself to win either way.
2026-08-25 ~ 2026-08-26 · 4 episodes · 55 posts
Episode 1 · OpenAI's First Custom Inference Chip Jalapeño Beats Nvidia in Tests (2026-08-25, 48 posts)
OpenAI has released benchmark results for Jalapeño, its first in-house inference chip, claiming it beats NVIDIA's flagship GB200 and GB300 systems across multiple inference workloads. The chip was co-developed with Broadcom, going from team formation to tape-out in only about 16 months. The conclusions currently come from OpenAI's own InferenceX tests and an independent teardown analysis by SemiAnalysis, and are drawing wide attention because they could shake NVIDIA's position in the AI inference compute market.
Confirmed
- OpenAI unveiled Jalapeño, its first self-developed inference chip, co-developed with Broadcom, with a design cycle of only about 16 months (@dylan522p, @zephyrz9).
- On OpenAI's public InferenceX benchmark, tested with open-source models including GPT-OSS 120B, DeepSeek R1 670B, and Kimi K2, Jalapeño delivers higher peak throughput per kilowatt and lower token latency, with better TCO and throughput than NVIDIA (@m2, @m4, @m6).- At the same DeepSeek R1 decode speed, throughput per kilowatt is 104.3x that of NVIDIA GB300 (12,258 vs 118 tok/s/kW) (@rohanpaulai).
- Energy efficiency reaches up to 1.9x that of Nvidia GB200/GB300 (@kimmonismus).
- SemiAnalysis's deep-dive teardown notes that Jalapeño is not narrowly optimized only for OpenAI models but is a general-purpose LLM inference ASIC, with measured efficiency exceeding NVIDIA and AMD (@zephyrz9, @dylan522p).
- OpenAI says that versus NVIDIA Blackwell, Jalapeño leads on two categories of metrics including AI processing per unit of power (official statements relayed by @dinabass).
- The chip is positioned to improve model inference efficiency and lower compute costs (@Polymarket).
Not yet confirmed
- All performance figures come from OpenAI's own InferenceX benchmark and SemiAnalysis's analysis; there is no independent verification from NVIDIA or any authoritative third party yet. The test conditions behind extreme numbers like the "104.3x" claim (e.g., comparison system configurations, quantization, and batching settings) still await more public detail.
Why it matters
- This is OpenAI's first public disclosure of real-world data for its in-house chip, marking the landing phase of its strategy to reduce dependence on NVIDIA. If the general-purpose inference ASIC positioning holds up, it could put real competitive pressure on NVIDIA and AMD in the data center inference market.
- SemiAnalysis: OpenAI's Jalapeño Is a General-Purpose Inference ASIC Beating NVIDIA and AMD in Tests — zephyr_z9 · 2026-08-25
- OpenAI Reveals Jalapeño Benchmarks: Beats Blackwell in Efficiency — dinabass · 2026-08-25
- Media Reports on OpenAI Jalapeño Benchmark Data — shiringhaffary · 2026-08-25
- OpenAI's 'Jalapeño' Chip Beats Nvidia Blackwell in Benchmarks — dylan522p · 2026-08-25
- OpenAI claims custom 'Jalapeño' chip outperforms Nvidia GB200/GB300 in inference — Polymarket · 2026-08-25
- OpenAI's Jalapeño chip claims 104.3x better efficiency than Nvidia GB300 — rohanpaul_ai · 2026-08-25
- OpenAI's first custom inference chip Jalapeño claims 1.5-1.9x better perf-per-watt than Nvidia GB200/GB300 — kimmonismus · 2026-08-25
- Jalapeño ASIC outperforms comparable TPUs, challenging existing giants — GavinSBaker · 2026-08-25
- OpenAI Claims New 'Jalapeño' Chip Outperforms Vera Rubin in Benchmarks — Wonderful_Buffalo_32 · 2026-08-25
- OpenAI's First Custom Inference Chip Jalapeño Beats Commercial Systems — firstadopter · 2026-08-25
- Jalapeño ASIC outperforms comparable TPU in debut — GavinSBaker · 2026-08-25
- OpenAI's new Jalapeno chip beats Nvidia GB300 in efficiency — thesaraharminta · 2026-08-25
- Jalapeño ASIC benchmarks before TPU support, showing competitive performance — GavinSBaker · 2026-08-25
- Jalapeño Claims Industry-Leading AI Inference Speed and Efficiency in First Results — pstAsiatech · 2026-08-25
- OpenAI's in-house inference chip reportedly rivals GB300, NVIDIA impact seen as limited — ivan_bezdomny · 2026-08-25
- OpenAI's first custom inference chip Jalapeño beats commercial systems in efficiency — firstadopter · 2026-08-25
- Deep Technical Overview of OpenAI's Upcoming Inference Chip: Jalapeno — Much_Preparation_832 · 2026-08-25
- Discussion: OpenAI's In-House Chip 'Jalapeño' May Outperform Nvidia Blackwell — thehiphopswami · 2026-08-25
- OpenAI Reveals Test Results for Custom Inference Chip Jalapeño — OpenAI · 2026-08-26
- OpenAI: Jalapeño Chip Speeds Up ChatGPT Responses and Codex — OpenAI · 2026-08-26
Episode 2 · OpenAI to Deploy Its Jalapeño Chip by Year-End with Next Generations in Development (2026-08-26, 2 posts)
OpenAI plans to deploy its custom Jalapeño chip by year-end as the first step in a multi-generation roadmap, with unconfirmed reports of major internal performance gains. The second-generation chip is in deep development and a third is taking shape.
- OpenAI to deploy Jalapeño chip by year-end, Gen 2 in deep development — OpenAI · 2026-08-26
- Rumor: OpenAI's Custom Chip Jalapeño Shows Big Gains, Deployment by Year-End — soumitrashukla9 · 2026-08-26
Episode 3 · OpenAI's Jalapeño Chip Play: Winning Either Way (2026-08-26, 2 posts)
SemiAnalysis analysis of OpenAI's custom ASIC, codenamed Jalapeño, suggests OpenAI wins regardless of deployment outcome, potentially extracting billions from Nvidia deals. Jalapeño reportedly outperforms Cerebras on cost and power efficiency.
- Deep Dive: OpenAI's Custom Chip Jalapeño May Outperform Cerebras — bookwormengr · 2026-08-26
- OpenAI's Chip Strategy: Winning Billions from Nvidia Even If the Chip Fails — AccBalanced · 2026-08-26
Episode 4 · OpenAI's In-House Chip Reportedly Beats Nvidia Blackwell on Efficiency (2026-08-26, 3 posts)
Leaks and SemiAnalysis testing indicate OpenAI's custom ASIC Jalapeño, with a 700W TDP, roughly doubles Nvidia Blackwell's throughput per watt in most no-finetuning inference scenarios, going from RTL to tape-out in just 9 months.
- OpenAI's Jalapeño Chip Reportedly Beats Nvidia Blackwell on Perf/Watt — petrusenko_max · 2026-08-26
- OpenAI's 'Jalapeno' chip reportedly beats Nvidia Blackwell in efficiency and latency benchmarks — 量子位 · 2026-08-26
- OpenAI's custom chip 'Jalapeño' reportedly beats Nvidia Blackwell in efficiency — yacineMTB · 2026-08-26