OpenAI says its Jalapeño chip beats Nvidia's GB300 on inference

OpenAI has published its first benchmark results for Jalapeño, its in-house inference chip built with Broadcom.
The company claims that the chip does more AI work per watt and returns answers faster than the Nvidia GB200 and GB300 rack systems that it was measured against.
How does OpenAI’s new chip perform?
OpenAI has published benchmark results for its new in-house AI chip, called Jalapeño. The chip was run through InferenceX, a public benchmark from SemiAnalysis that measures the full job of serving an AI request, and tested it on OpenAI’s own GPT-OSS 120B, DeepSeek’s R1 670B, and Moonshot AI’s Kimi K2.5 1T.
OpenAI said its new chip did 1.5 to 1.9 times more work for each unit of electricity than the other systems it was tested against. It also answered requests 1.7 to 3.6 times faster.
For the most hands-on tasks that require much back-and-forth, the gap grew to between 2.1 and 4.1 times better.
OpenAI’s hardware vice president, Richard Ho, told reporters the chip handles more work at once and also responds quicker, which most chips cannot do at the same time.
… Continue reading the full article at the original source below.



