
OpenAI says its new AI chip, Jalapeño, completes tasks more efficiently and returns responses faster than other AI systems, according to a blog post published on Tuesday. During a briefing with reporters, OpenAI hardware vice president Richard Ho said Jalapeño offers the "best of both worlds" with lower latency and higher throughput, as AI systems typically "have to make a trade-off between the two." First introduced in June, Jalapeño is an Application-Spec
OpenAI has developed a new AI chip called Jalapeño, an Application-Specific Integrated Circuit made in partnership with Broadcom, designed specifically to run trained AI models to complete tasks. According to OpenAI's testing, Jalapeño outperforms competitor chips by delivering responses faster while using less energy, achieving 1.5 to 1.9 times more AI work per watt and 1.7 to 3.6 times lower response time compared to Nvidia's superchips. The company plans to begin deploying Jalapeño in limited quantities by year's end and increase production into the following year, though it will continue using chips from other partners like Nvidia rather than replacing its entire chip lineup. This development is part of an ongoing competition among major technology companies to develop custom AI chips for improved performance and efficiency.

Model cards report quality under server-class, full-precision conditions. Those numbers rarely predict how the same model behaves on a phone. This week, Liquid AI released Pipette. It is an open-source platform for benchmarking foundation models on edge devices, built in partnership with Artificial Analysis as an independent methodology validator. Pipette treats on-device behavior as a property of the deployed system, not the model in isolation. Its unit of measurement is a full configuration:

Tested on SemiAnalysis’ InferenceX benchmark, Jalapeño registered both more tokens per user and more throughput per kilowatt than the currently available state-of-the art.

Shrinking neural networks to use less memory and compute while maintaining or improving accuracy could make AI models cheaper to deploy widely.
Want to go deeper than the news? Explore live, cohort-based AI courses taught by practitioners.
Browse AI courses on Maven