OpenAI's first custom inference chip Jalapeño beats commercial systems in efficiency
firstadopter · x · 2026-08-25
OpenAI released the first measured performance results for Jalapeño, its custom inference chip. On the InferenceX benchmark using GPT‑OSS 120B, Jalapeño delivered higher peak throughput per kilowatt and lower token latency than commercial systems. It also performed strongly on DeepSeek R1 and Kimi K2, showing cross-family adaptability. OpenAI highlights this as an example of Jevons paradox: greater efficiency expands consumption and creates new economic activity.
More from Infra
- OpenAI Used Unreleased Model Astra to Design Its Jalapeño Inference Chip in 9 Months — jarrodwatts · 2026-08-26
- NVIDIA launches Jetson Orin Nano 2: 2x performance, 40% less power — nvidia · 2026-08-26
- Studies Reveal Data Centers Boost Local Jobs and Wages Significantly — justin_hart · 2026-08-26
- Opinion: Land Scarcity Drives Shift to Decentralized AI Compute — bittingthembits · 2026-08-26
- Jalapeno beats VR200 with optimized DeepSeek implementation, faster execution — itsclivetime · 2026-08-26
- CUDA code now runs on Apple Silicon with zero source code changes — petewoodbridge · 2026-08-26