OpenAI's Jalapeño ASIC claims faster inference, lower latency than Nvidia chips
OpenAI announced that its Jalapeño ASIC, developed with Broadcom, delivers 1.5x-1.9x more work per watt and 1.7x-3.6x lower latency than Nvidia chips on inference benchmarks across models like GPT-OSS, DeepSeek R1, and Kimi K2.5 1T. The chip, introduced in June, is designed for AI inference and aims to balance latency and throughput. SemiAnalysis reported that Jalapeño outperformed Nvidia, AMD, and Google chips on multiple open-weight models, and was developed in 16 months.
Coverage timeline
The Verge AIEmma Roth
OpenAI says its new AI chip, Jalapeño, completes tasks more efficiently and returns responses faster than other AI systems, according to a blog post published on Tuesday. During a briefing with reporters, OpenAI hardware vice president Richard Ho said Jalapeño offers the "best of both worlds" with lower latency and higher throughput, as AI systems typically "have to make a trade-off between the two." First introduced in June , Jalapeño is an Application-Specific Integrated Circuit (ASIC) made in partnership with Broadcom. It's designed for AI inference - the process of running a trained AI model to complete a task or deploy an agent. To mea … Read the full story at The Verge.

Techmeme
Emma Roth / The Verge : OpenAI says its Jalapeño chip delivered 1.5x-1.9x more AI work per watt and 1.7x-3.6x lower latency vs. Nvidia chips across GPT-OSS, DeepSeek R1, Kimi K2.5 1T — Jalapeño outperformed Nvidia's superchips on an AI inference benchmark test.

Techmeme
SemiAnalysis : A detailed look at Jalapeño, OpenAI's ASIC developed with Broadcom in 16 months, which beat Nvidia, AMD, and Google chips on multiple top open-weight models — OpenAI's self-designed ASIC compared with Rubin, Jalapeño's TCO, throughput per MW, and spicy deets
