OpenAI has unveiled results from testing Jalapeno, its first custom AI inference chip, claiming major gains in speed, latency and power efficiency. Tested across GPT-OSS 120B, DeepSeek R1 and Kimi K2.5 1T, the chip delivered up to 1.9x more work per watt and 3.6x lower latency. OpenAI plans to deploy Jalapeno in its compute infrastructure by year-end.
California: OpenAI has announced results from testing Jalapeno, its first custom inference chip, saying the in-house silicon delivers higher throughput, lower latency and greater power efficiency across multiple AI models.
In a post on X, OpenAI CEO Sam Altman announced the development, saying, "we made a chip and it is fast".
According to an article titled 'Jalapeno's first results show industry-leading speed and efficiency in AI inference' published on OpenAI's official website, on Tuesday (local time) the company said Jalapeno can "serve more AI work per unit of power while also returning responses more quickly".
"Jalapeno delivers both higher throughput and lower latency with one architecture, where existing hardware systems often have to make a tradeoff between the two," OpenAI said in the article.
-
GTA 6 previews achieve 31.1 million views on Netflix

-
Does wearing a hat continuously cause hair fall? Know what to do if it is so…

-
Your Love Horoscope Is Here For Thursday, September 3, 2026

-
Tamarind isn’t just a cooking ingredient: 5 ways it can help make kitchen cleaning easier |

-
Tamarind isn’t just a cooking ingredient: 5 ways it can help make kitchen cleaning easier |
