OpenAI Jalapeño Chip vs Nvidia: 3.6x Lower Latency [2026]

OpenAI Jalapeño Chip vs Nvidia: 3.6x Lower Latency [2026]

By shattered.io
Publication Date: 2026-10-04 21:11:00

OpenAI put real numbers behind its custom silicon bet this week, and the figures land squarely in Nvidia’s backyard. The company’s first in-house inference chip, codenamed Jalapeño, posted benchmark results showing 1.5 to 1.9 times more AI work per watt and 1.7 to 3.6 times lower end-to-end latency than the comparison systems OpenAI tested against, according to the company’s own August 25, 2026 results post. For a market that has run almost entirely on Nvidia GPUs since the generative AI boom began, that is a loaded claim, and it is the clearest evidence yet that OpenAI intends to build, not just buy, the chips that run ChatGPT.

Jalapeño was first announced in partnership with Broadcom on June 24, 2026, as OpenAI’s first custom inference chip, what the company calls an “Intelligence Processor.” The design goal…