OpenAI says its Jalapeño chip outperforms Nvidia’s superchips in AI inference benchmarks
OpenAI's Jalapeño chip outperforms Nvidia's superchips in AI inference benchmarks, delivering 1.5 to 1.9 times more AI work per watt and 1.7 to 3.6 times lower end-to-end latency.
Intelligence analysis by Qwen 2.5 (3B)

OpenAI's Jalapeño chip, designed for AI inference, outperforms Nvidia's superchips in benchmarks, offering better efficiency and faster responses.
OpenAI's new chip, Jalapeño, can do more work with less energy and make AI systems faster and more reliable, like how a faster computer can do more tasks in less time.
Analysis
{"
Jalapeño Chip Performance Details":"Jalapeño outperformed Nvidia's superchips on an AI inference benchmark test, delivering 1.5 to 1.9 times more AI work per watt and 1.7 to 3.6 times lower end-to-end latency across three models.","
OpenAI's Strategy":"OpenAI plans to deploy Jalapeño in small volumes by the end of the year, with plans to ramp up volume in 2027. The company expects to continue developing the chip's second and third generations.","
Industry Context":"This development comes as other tech giants like Microsoft, Meta, Google, and Nvidia are also developing AI chips. OpenAI's Jalapeño chip aims to offer a balance between lower latency and higher throughput, addressing the trade-off typically faced by AI systems."}
Key points
- Jalapeño outperforms Nvidia's superchips in AI inference benchmarks
- OpenAI plans to deploy Jalapeño in small volumes by the end of the year
- Jalapeño aims to offer a balance between lower latency and higher throughput
Jalapeño could lead to more efficient AI systems, potentially reducing energy costs and improving user experience.
However, OpenAI might not replace its entire chip lineup with Jalapeño, as it plans to continue working with existing partners like Nvidia.



