AI Chips4 mins read

OpenAI’s Jalapeño Chip Reportedly Tops Nvidia Blackwell and Rubin in AI Inference Tests

OpenAI’s first in-house inference chip, Jalapeño, reportedly outperforms Nvidia’s Blackwell and Rubin systems in key inference benchmarks, according to results discussed at Hot Chips and analyzed by SemiAnalysis.

OpenAI Jalapeño chip image
Image credits:OpenAI

What OpenAI Claimed at Hot Chips

OpenAI showed benchmarks for Jalapeño, its first in-house inference chip, at the Hot Chips conference. The chip is built for inference only, meaning it runs AI models rather than training them. The Decoder reports that Jalapeño is positioned as a general-purpose LLM inference accelerator rather than a chip tuned only for OpenAI’s own models.

Where Jalapeño Reportedly Beats Nvidia

Benchmark chart comparing Jalapeño token throughput per kilowatt with other accelerators
Image credits:OpenAI

According to SemiAnalysis testing cited by The Decoder, Jalapeño beats Nvidia’s Blackwell and even Rubin in throughput and energy efficiency. OpenAI claims the chip delivers 1.5x to 1.9x more AI work per watt at peak throughput across three tested models. The company also claims 1.7x to 3.6x lower end-to-end latency than the best commercially available systems, and 2.1x to 4.1x higher performance for interactive workloads.

Benchmark Details and Key Caveats

The tests used SemiAnalysis’s public InferenceX benchmark, with OpenAI providing the numbers and SemiAnalysis verifying some runs on-site in the lab. The tested models were GPT-OSS 120B, Deepseek R1 670B, and Kimi K2.5 1T; Jalapeño reportedly reached about 1,400 tokens per second per user on GPT-OSS and more than 700 tokens per second on a single concurrent Deepseek R1 request. The Decoder notes important limits: Nvidia and AMD have published results on larger models not yet tested on Jalapeño, and Rubin systems are already shipping while Jalapeño reportedly remains at the engineering-sample stage.

Why This Matters for the AI Chip Race

OpenAI developed Jalapeño with Broadcom, and the company says it used its own AI models during development. The reported speed of development and the benchmark results raise questions about how durable Nvidia’s software and hardware advantage remains. OpenAI CFO Sarah Friar frames the chip as part of a broader compute strategy that complements partnerships with Nvidia, AMD, AWS, Cerebras, CoreWeave, and others rather than replacing them.

Discover More