OpenAI to Deploy Custom Jalapeño AI Inference Chip in Its Data Centers by 2026
OpenAI plans to deploy Jalapeño, its first purpose-built chip for large language model inference, within its own compute infrastructure by the end of 2026. Developed in collaboration with Broadcom, the processor is designed to handle workloads powering ChatGPT, Codex, and OpenAI's API rather than being sold directly to customers. The chip went from design to manufacturing tape-out in nine months, and OpenAI claims early testing shows significantly better performance per watt than current leading accelerators. OpenAI is also developing second and third generations of the chip as part of a multi-generation roadmap targeting improved efficiency and speed. The move is part of a broader full-stack infrastructure effort spanning hardware, software, networking, and large-scale data center deployments.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in