OpenAI's Jalapeño chip outperforms NVIDIA on speed, throughput, and power efficiency
OpenAI has published benchmark results for Jalapeño, its first custom AI inference chip, showing it outperforms NVIDIA's GB200 and GB300 hardware across throughput, latency, and power efficiency simultaneously. Tested against the InferenceX public benchmark on three large open-weight models, Jalapeño delivered 1.5–1.9x more throughput per watt and 1.7–3.6x lower latency compared to leading NVIDIA hardware. The chip operates at or below 550W, roughly half the power draw of the GB300 at 1,400W. Designed in nine months, Jalapeño reflects OpenAI's broader strategy to co-design models, chips, and software together rather than rely on third-party GPU infrastructure. OpenAI plans to deploy the chip within its own infrastructure by end of 2025, with a second generation already in development.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in