OpenAI and Broadcom have announced Jalapeño, a custom ASIC designed for large language model inference in data centers. The chip promises substantially better performance per watt than current state-of-the-art.
<p><!-- obsidian --></p>
<p>OpenAI, the company behind ChatGPT and Codex and the models those tools utilize, and Broadcom, an established silicon supplier, have <a href="https://openai.com/index/openai-broadcom-jalapeno-inference-chip/">announced</a> a new chip called Jalapeño, designed specifically for large language model inference in data centers.</p>
<p>The chip is intended to be deployed at large data centers, both companies claim this is just the first generation in a long-term project that will see chips refined over time.</p><p><a href="https://arstechnica.com/gadgets/2026/06/openai-and-broadcom-announce-chip-designed-for-llm-inference-at-scale/">Read full article</a></p>
<p><a href="https://arstechnica.com/gadgets/2026/06/openai-and-broadcom-announce-chip-designed-for-llm-inference-at-scale/#comments">Comments</a></p>
# OpenAI and Broadcom announce chip designed for LLM inference at scale
Source: [https://arstechnica.com/gadgets/2026/06/openai-and-broadcom-announce-chip-designed-for-llm-inference-at-scale/](https://arstechnica.com/gadgets/2026/06/openai-and-broadcom-announce-chip-designed-for-llm-inference-at-scale/)
OpenAI, the company behind ChatGPT and Codex and the models those tools utilize, and Broadcom, an established silicon supplier, have[announced](https://openai.com/index/openai-broadcom-jalapeno-inference-chip/)a new chip called Jalapeño, designed specifically for large language model inference in data centers\.
The chip is intended to be deployed at large data centers, both companies claim this is just the first generation in a long\-term project that will see chips refined over time\.
Broadcom says that this ASIC \(Application\-Specific Integrated Circuit\) was designed from scratch for LLM inference, based on “detailed insights” from the company’s conversations with researchers at OpenAI, and that the chip’s development was informed by OpenAI’s own roadmap for future models and products\. The design and production of the chip took nine months\.
The promise is that this chip is more specialized for the current needs of LLMs than those that inference systems currently run on in existing data centers\.
OpenAI claims that “early testing shows that Jalapeño will deliver performance per watt substantially better than current state\-of\-the\-art,” but notes that it is not done measuring performance, and that a “detailed technical report will be presented in the coming months\.”
OpenAI and Broadcom unveiled Jalapeño, a custom LLM-optimized inference chip that promises substantially better performance per watt than current state-of-the-art, designed from the ground up for current and future AI models.
OpenAI's Jalapeño chip, developed with Broadcom, shows significant performance gains in AI inference benchmarks, surpassing state-of-the-art processors like Nvidia's Blackwell. The chip is optimized for fast, efficient inference at scale, with deployment starting in late 2026.
OpenAI unveiled its first custom-built inference processor, named Jalapeño, developed with Broadcom to improve performance-per-watt and reduce reliance on Nvidia GPUs.
OpenAI has announced its first custom AI inference chip, Jalapeño, developed in partnership with Broadcom to reduce reliance on Nvidia GPUs, with deployment expected by the end of 2026.
OpenAI's Jalapeño custom inference chip shows industry-leading speed and efficiency, delivering higher throughput and lower latency across various AI models.