Top News

'We Made A Chip And It Is Fast': OpenAI Unveils Jalapeno AI Inference Chip With Major Gains In Speed, Latency And Power Efficiency
24htopnews | August 26, 2026 4:08 PM CST

OpenAI has unveiled results from testing Jalapeno, its first custom AI inference chip, claiming major gains in speed, latency and power efficiency. Tested across GPT-OSS 120B, DeepSeek R1 and Kimi K2.5 1T, the chip delivered up to 1.9x more work per watt and 3.6x lower latency. OpenAI plans to deploy Jalapeno in its compute infrastructure by year-end.

California: OpenAI has announced results from testing Jalapeno, its first custom inference chip, saying the in-house silicon delivers higher throughput, lower latency and greater power efficiency across multiple AI models.

In a post on X, OpenAI CEO Sam Altman announced the development, saying, "we made a chip and it is fast".

According to an article titled 'Jalapeno's first results show industry-leading speed and efficiency in AI inference' published on OpenAI's official website, on Tuesday (local time) the company said Jalapeno can "serve more AI work per unit of power while also returning responses more quickly".

"Jalapeno delivers both higher throughput and lower latency with one architecture, where existing hardware systems often have to make a tradeoff between the two," OpenAI said in the article.


READ NEXT
Cancel OK