# Groq Licenses LPU Architecture to Nvidia in $20B Deal

> In a surprising twist, Nvidia is licensing technology from a smaller rival. It paid roughly $20 billion for Groq's LPU inference architecture, a chip design built for one job: generating tokens as fast as possible. The Groq 3 LPU reaches about 1,500 tokens per second, leaning on 500MB of on-chip SRAM and 150 terabytes per second of internal bandwidth. The trade-offs are deliberate. An LPU gives up trainability and vision work in exchange for raw memory bandwidth and blistering autoregressive generation, which is exactly the profile that agentic AI, with its long chains of tool calls and reasoning steps, is hungry for. Nvidia plans to pair the design with its Vera Rubin systems in five new rack configurations shown at GTC 2026. Groq founders Jonathan Ross and Sunny Madra are moving to Nvidia, though Groq continues independently under new leadership. The deal is a tell about where inference is heading: as agents multiply, the speed and cost of producing tokens, not just training bigger models, becomes the battleground.

_Section: [Daily AI Updates](https://www.wortins.com/daily-ai) · Source: Spheron · Published Thursday, July 16, 2026_

## Wortins' read

In a surprising twist, Nvidia is licensing technology from a smaller rival. It paid roughly $20 billion for Groq's LPU inference architecture, a chip design built for one job: generating tokens as fast as possible. The Groq 3 LPU reaches about 1,500 tokens per second, leaning on 500MB of on-chip SRAM and 150 terabytes per second of internal bandwidth. The trade-offs are deliberate. An LPU gives up trainability and vision work in exchange for raw memory bandwidth and blistering autoregressive generation, which is exactly the profile that agentic AI, with its long chains of tool calls and reasoning steps, is hungry for. Nvidia plans to pair the design with its Vera Rubin systems in five new rack configurations shown at GTC 2026. Groq founders Jonathan Ross and Sunny Madra are moving to Nvidia, though Groq continues independently under new leadership. The deal is a tell about where inference is heading: as agents multiply, the speed and cost of producing tokens, not just training bigger models, becomes the battleground.

## Source

[Read the full story at Spheron](https://www.spheron.network/blog/nvidia-groq-3-lpu-explained/)

## Related coverage

- [Top AI spenders cut per-employee costs by nearly 10 percent in August](https://www.wortins.com/story/top-ai-spenders-cut-per-employee-costs-by-nearly-10-percent--b31ffb0b) — [The Decoder](https://the-decoder.com/top-ai-spenders-cut-per-employee-costs-by-nearly-10-percent-in-august/)
- [A look at why the oft-discussed predictions that AI will deliver double-digit GDP growth in advanced economies are extremely unlikely over the next 10-15 years (Ghosts of Electricity)](https://www.wortins.com/story/a-look-at-why-the-oft-discussed-predictions-that-ai-will-del-6241fa38) — [Techmeme](https://www.techmeme.com/260910/p10#a260910p10)
- [Farmers Embrace AI More Than Any Other Tech, McKinsey Says](https://www.wortins.com/story/farmers-embrace-ai-more-than-any-other-tech-mckinsey-says-78dfb3ec) — [Insurance Journal](https://www.insurancejournal.com/news/national/2026/09/09/884469.htm)
- [Microsoft has new AI privacy rules for schools](https://www.wortins.com/story/microsoft-has-new-ai-privacy-rules-for-schools-6e0b8ab0) — [The Verge](https://www.theverge.com/policy/992359/microsoft-aft-schools-ai-privacy)
- [Chinese tech giants are hiring skilled professionals as specialized AI trainers to build high-quality datasets, mirroring efforts by US platforms like Mercor (Viola Zhou/Rest of World)](https://www.wortins.com/story/chinese-tech-giants-are-hiring-skilled-professionals-as-spec-116fb4b8) — [Techmeme](https://www.techmeme.com/260910/p6#a260910p6)
- [Researchers used AI to build a WeChat worm that spreads through phone calls](https://www.wortins.com/story/researchers-used-ai-to-build-a-wechat-worm-that-spreads-thro-8bf3f469) — [The Next Web](https://thenextweb.com/news/wechat-worm-ai-calif-tencent-zero-click)

---

_Curated and written by [Wortins](https://www.wortins.com) — The daily AI briefing. Every story links to its original source; the "Wortins read" on each is our own original analysis. [About Wortins & our editorial approach](https://www.wortins.com/about)._
