OpenAI Shows First Jalapeño Chip Results
OpenAI published the first results for Jalapeño, its custom inference chip, while also shipping new ChatGPT browser and scheduled task features. On SemiAnalysis’s public InferenceX benchmark, Jalapeño delivered 1.5 to 1.9 times more AI work per watt and 1.7 to 3.6 times lower latency than Nvidia GB200 and GB300 systems, with SemiAnalysis saying it beats every Nvidia, AMD, and Google chip it has tested.
- OpenAI says Jalapeño’s performance held across three public model families: GPT-OSS 120B, DeepSeek R1 670B, and Kimi K2.5 1T.
- The chip is rated at 700 watts, while measured sustained power stayed at or below 550 watts on the tested workloads.
- OpenAI says AI helped move the chip from initial design to tapeout in nine months, and AI-generated implementations beat selected human-written blocks by 1.5 to 1.8 times.
