OpenAILaunchAI ModelsAugust 25, 2026

OpenAI's Jalapeño chip beats Nvidia in inference benchmarks

OpenAI's custom inference chip Jalapeño delivered 1.5–1.9× more AI work per watt and 1.7–3.6× lower latency than Nvidia GB200/GB300 on InferenceX across GPT-OSS 120B, DeepSeek R1, and Kimi K2.5 1T. Deployment starts in small volumes by end of 2026, ramping in 2027.

Featured · Richard Ho

15 sources

OpenAI by email

Get an email when OpenAI has news

No news that day, no email.

More stories today

Open the live feed