OpenAI's Jalapeño chip beats Nvidia in inference benchmarks

OpenAI's custom inference chip Jalapeño delivered 1.5–1.9× more AI work per watt and 1.7–3.6× lower latency than Nvidia GB200/GB300 on InferenceX across GPT-OSS 120B, DeepSeek R1, and Kimi K2.5 1T. Deployment starts in small volumes by end of 2026, ramping in 2027.
Featured · Richard Ho
15 sources
Jalapeño’s first results show industry-leading speed and efficiency in AI inferenceopenai.com
we made a chip and it is fastx.com
OpenAI Claims Its New Chips Can Outperform Nvidia Processors in Testsbloomberg.com
OpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks showtechcrunch.com
OpenAI says its Jalapeño chip can power faster AI responses than the competitiontheverge.com
OpenAI built a chip in nine months. Then it let AI rewrite the code.thenewstack.io
OpenAI Just Dropped Benchmarks for Their Own Chip, Jalapeño, and It's Beating Nvidia's GB300reddit.com
OpenAI JalapeñO: Better Than Nvidia Blackwellnewsletter.semianalysis.com
OpenAI by email
Get an email when OpenAI has news
No news that day, no email.
More stories today
- Liquid AI open-sources Pipette benchmarking suite for on-device models
- wikiHow sues OpenAI over copyright infringement in AI training
- Claude Code 2.1.246 adds Auto mode tab, Bash wildcard warning
- Korean AI startup Wrtn raises funds at $870M valuation
- Podcast: Google DeepMind's Vivek Natarajan on AI in healthcare