OpenAI's Jalapeño chip delivers fast, scalable inference, benchmarks reveal
benchmarks chips inference openai
| Source: TechCrunch | Original article
OpenAI's new Jalapeño chip delivers higher token output per user and greater throughput per kilowatt than existing state‑of‑the‑art hardware, according to SemiAnalysis' InferenceX benchmark.
OpenAI’s Jalapeño inference chip has posted its first independent benchmark results, confirming the performance edge the company hinted at in its August announcements. Testing on SemiAnalysis’s InferenceX suite, Jalapeño delivered more tokens per user and higher throughput per kilowatt than any currently available inference processor. Across a range of models—including GPT‑OSS‑120B and DeepSeek—the chip achieved between 1.5 × and 1.9 × more “AI work” at peak throughput while cutting end‑to‑end latency by roughly 1.7 × to 3.6 × relative to the competition. The results also show Jalapeño outpacing Nvidia’s Blackwell platform on the same metrics.
The breakthrough matters because inference efficiency directly translates into operating costs and environmental impact for the massive data‑center fleets that power chat‑bots, recommendation engines and other AI services. Higher token output per user means providers can serve more queries without adding hardware, while the superior power‑per‑throughput ratio eases the growing electricity demands of AI workloads. If the chip lives up to its lab numbers at scale, it could reshape the economics of AI serving and pressure rivals to accelerate their own efficiency roadmaps.
OpenAI plans a very small‑scale deployment of Jalapeño by the end of 2026, followed by a broader rollout in 2027. The next steps will be watching how the chip integrates into OpenAI’s own API infrastructure and whether cloud operators adopt it for third‑party workloads. Equally important will be further independent benchmarks and real‑world usage data that confirm whether Jalapeño can sustain its lead over Nvidia’s Blackwell and other emerging inference solutions. As we reported on Aug 26, OpenAI’s chip strategy is already challenging Nvidia’s dominance; these fresh results tighten the race.
Sources
Back to AIPULSEN