OpenAI’s Jalapeño Beats Nvidia’s Blackwell and Rubin in Tests #
OpenAI’s first in-house inference chip delivered 1.5 to 1.9 times more AI work per watt and 1.7 to 3.6 times lower end-to-end latency than comparison systems. SemiAnalysis tests presented at the Hot Chips conference found Jalapeño ahead of Nvidia’s Blackwell and Rubin in throughput and energy efficiency.
On SemiAnalysis’ InferenceX benchmark, Jalapeño produced more tokens per user and more throughput per kilowatt than the currently available state of the art. For highly interactive workloads, it delivered 2.1 to 4.1 times higher performance.
OpenAI unveiled the chip program in June with Broadcom, building Jalapeño from a blank sheet for large-scale inference.
Why it matters: OpenAI now has benchmark evidence that its in-house inference chip outperforms Nvidia’s Blackwell and Rubin on energy efficiency, latency and interactive-workload performance.
Key Takeaways
- OpenAI showed Jalapeño at the Hot Chips conference, where the benchmark results were presented
- OpenAI and Broadcom unveiled the program in June and built the chip from a blank sheet
- SemiAnalysis’ InferenceX test measured more tokens per user and more throughput per kilowatt for Jalapeño