# OpenAI's Jalapeño Inference Chip Beats Nvidia Blackwell on Performance Per Watt

> OpenAI walked into Hot Chips on August 25 with benchmarks for Jalapeño, its first serious custom inference chip, and the numbers were aimed straight at Nvidia. On interactive chatbot-style workloads the company claimed 1.5 to 1.9 times better energy efficiency and 2.1 to 4.1 times lower latency than Nvidia's Blackwell, the part that currently underpins much of the industry. The caveats matter. Jalapeño only handles inference, the running of already-trained models, and leaves training firmly in Nvidia's court. It was built with Broadcom and is only headed for low-volume production late this year, so it will not dent Nvidia's order book any time soon. Still, the direction is clear. Every large model company is now trying to escape the Nvidia tax on the workload that actually costs them money day to day, which is serving billions of chat responses. If OpenAI can really cut the power bill on inference by a third or more, it changes the economics of running its own products and gives it leverage it has never had over its most important supplier.

_Section: [Daily AI Updates](https://www.wortins.com/daily-ai) · Source: CNBC · Published Thursday, August 27, 2026_

## Wortins' read

OpenAI walked into Hot Chips on August 25 with benchmarks for Jalapeño, its first serious custom inference chip, and the numbers were aimed straight at Nvidia. On interactive chatbot-style workloads the company claimed 1.5 to 1.9 times better energy efficiency and 2.1 to 4.1 times lower latency than Nvidia's Blackwell, the part that currently underpins much of the industry. The caveats matter. Jalapeño only handles inference, the running of already-trained models, and leaves training firmly in Nvidia's court. It was built with Broadcom and is only headed for low-volume production late this year, so it will not dent Nvidia's order book any time soon. Still, the direction is clear. Every large model company is now trying to escape the Nvidia tax on the workload that actually costs them money day to day, which is serving billions of chat responses. If OpenAI can really cut the power bill on inference by a third or more, it changes the economics of running its own products and gives it leverage it has never had over its most important supplier.

## Source

[Read the full story at CNBC](https://www.cnbc.com/2026/08/26/openai-jalapeno-ai-chip-nvidia.html)

## Related coverage

- [Alibaba Raises $10.2 Billion in Record Hong Kong Share Sale to Fund AI Expansion](https://www.wortins.com/story/alibaba-raises-10-2-billion-in-record-hong-kong-share-sale-t-ee1afa0a) — [Bloomberg](https://www.bloomberg.com/news/articles/2026-08-23/alibaba-to-raise-10-billion-by-selling-shares-for-ai-expansion)
- [Moonshot AI Releases Kimi K3, World's Largest Open-Source AI Model at 2.8 Trillion Parameters](https://www.wortins.com/story/moonshot-ai-releases-kimi-k3-world-s-largest-open-source-ai--cde823ba) — [Tom's Hardware](https://www.tomshardware.com/tech-industry/artificial-intelligence/moonshot-releases-2-8-trillion-parameter-kimi-k3)
- [Checking In on AI and the Big Five](https://www.wortins.com/story/checking-in-on-ai-and-the-big-five-8f54332c) — [Stratechery](https://stratechery.com/2025/checking-in-on-ai-and-the-big-five/)
- [Cohere Launches Command A+ Mixture-of-Experts Model](https://www.wortins.com/story/cohere-launches-command-a-mixture-of-experts-model-5d840970) — [Cohere](https://docs.cohere.com/docs/command-a-plus)
- [SoftBank Plans Record $6.3 Billion Retail Bond Sale to Fund OpenAI Investment](https://www.wortins.com/story/softbank-plans-record-6-3-billion-retail-bond-sale-to-fund-o-dbc93b82) — [Bloomberg](https://www.bloomberg.com/news/videos/2026-08-20/bloomberg-tech-8-20-2026-video)
- [Emerald AI Raises $150 Million Series A at $1.05 Billion Valuation](https://www.wortins.com/story/emerald-ai-raises-150-million-series-a-at-1-05-billion-valua-2071899e) — [Business Wire](https://www.businesswire.com/news/home/20260825127649/en/Emerald-AI-Raises-$150-Million-Series-A-at-$1.05-Billion-Valuation-to-Scale-Power-Flexible-AI-Data-Centers)

---

_Curated and written by [Wortins](https://www.wortins.com) — The daily AI briefing. Every story links to its original source; the "Wortins read" on each is our own original analysis. [About Wortins & our editorial approach](https://www.wortins.com/about)._
