OpenAI's upcoming Jalapeño chip looks like it'll be an inference beast
128 chips, 1.7 exaFLOPS, and 27 TB of HBM give Altman and crew a leg up over Blackwell, and maybe even Rubin
OpenAI unveiled its upcoming Jalapeño AI accelerator chip at the Hot Chips semiconductor development conference in Stanford. This chip, developed in collaboration with Broadcom, is the first of a series of custom silicon from OpenAI designed for AI inference tasks. Compared to Nvidia's GPU systems, OpenAI claims Jalapeño will deliver higher throughput and lower latency when it becomes available later this year and reaches volume production in 2027.
Though Jalapeño won't replace OpenAI's existing hardware partners, it's designed with AI in mind. Memory bandwidth is crucial for inference, and Jalapeño appears to excel in this area. Early benchmarks show the chip delivering 1.5x to 1.9x more AI work at peak throughput and 1.7x to 3.6x lower end-to-end latency than competitors like GPT-OSS-120B, DeepSeek R1, and Kimi K2.5. For ultra-low-latency inference, Jalapeño is 2.1x to 4.1x faster.
Each Jalapeño system contains 128 accelerators, providing 1.7 exaFLOPS of 4-bit compute, 27.5 TB of HBM4 memory, and nearly 2 petabytes per second of memory bandwidth. Compared to AMD and Nvidia's latest rack systems, Jalapeño uses between 40 and 60 percent of the power. The system design is based on a rack scale architecture similar to Nvidia's NVL72 or AMD Helios.
Written by urgent.news from The Register Science's reporting — not their text. Machine-written — may contain errors; check the original before relying on it.
This story
This is one outlet's version. Read the fullest account.
- OpenAI's upcoming Jalapeño chip looks like it'll be an inference beast theregister.com
- OpenAI says its Jalapeno AI chip delivers faster responses than rivals like Nvidia digitaltrends.com
- OpenAI says its Jalapeño chip delivered 1.5x-1.9x more AI work per watt and 1.7x-3.6x lower latency vs. Nvidia chips across GPT-OSS, DeepSeek R1, Kimi K2.5 1T (Emma Roth/The Verge) theverge.com
- OpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks show techcrunch.com
- A detailed look at Jalapeño, OpenAI's ASIC developed with Broadcom in 16 months, which beat Nvidia, AMD, and Google chips on multiple top open-source models (SemiAnalysis) newsletter.semianalysis.com