# OpenAI Previews Ultrafast Mode for GPT-5.6 Sol: 14x Faster Processing Speed

> OpenAI and Cerebras Systems have unveiled Ultrafast, a new API tier that runs GPT-5.6 Sol at roughly 750 output tokens per second, about 14 times the speed of the standard mode. It is in limited preview only, with no pricing and no general availability date yet, so for now it is a demonstration of what the model can do when the hardware underneath it is built for raw throughput. The engine here is Cerebras, whose wafer-scale chips are designed to spit out tokens far faster than conventional GPU clusters. This tier is part of a broader OpenAI-Cerebras arrangement, a deal for 750 megawatts of compute capacity running through 2028 and valued at more than 10 billion dollars. Speed is quietly becoming its own frontier. As agents chain dozens of model calls together, latency compounds, and a 14x jump changes what feels possible in real time, from live coding to voice to interactive tools. It also signals that OpenAI is willing to look beyond Nvidia for the silicon that powers its fastest offerings.

_Section: [Daily AI Updates](https://www.wortins.com/daily-ai) · Source: TechCrunch · Published Sunday, August 23, 2026_

## Wortins' read

OpenAI and Cerebras Systems have unveiled Ultrafast, a new API tier that runs GPT-5.6 Sol at roughly 750 output tokens per second, about 14 times the speed of the standard mode. It is in limited preview only, with no pricing and no general availability date yet, so for now it is a demonstration of what the model can do when the hardware underneath it is built for raw throughput. The engine here is Cerebras, whose wafer-scale chips are designed to spit out tokens far faster than conventional GPU clusters. This tier is part of a broader OpenAI-Cerebras arrangement, a deal for 750 megawatts of compute capacity running through 2028 and valued at more than 10 billion dollars. Speed is quietly becoming its own frontier. As agents chain dozens of model calls together, latency compounds, and a 14x jump changes what feels possible in real time, from live coding to voice to interactive tools. It also signals that OpenAI is willing to look beyond Nvidia for the silicon that powers its fastest offerings.

## Source

[Read the full story at TechCrunch](https://techcrunch.com/2026/08/13/openai-introduces-ultrafast-a-new-mode-that-makes-gpt-5-6-sol-work-at-14x-the-speed/)

## Related coverage

- [Emerald AI Raises $150 Million Series A at $1.05 Billion Valuation](https://www.wortins.com/story/emerald-ai-raises-150-million-series-a-at-1-05-billion-valua-2071899e) — [Business Wire](https://www.businesswire.com/news/home/20260825127649/en/Emerald-AI-Raises-$150-Million-Series-A-at-$1.05-Billion-Valuation-to-Scale-Power-Flexible-AI-Data-Centers)
- [ShepHertz Launches AgentAnywhere, Sovereign Agentic AI Platform for Regulated Industries](https://www.wortins.com/story/shephertz-launches-agentanywhere-sovereign-agentic-ai-platfo-a80b8050) — [ShepHertz](https://aiagentstore.ai/ai-agent-news/this-week)
- [Apple Updates, AI Computers, and OpenAI's Jalapeño Chip](https://www.wortins.com/story/apple-updates-ai-computers-and-openai-s-jalape-o-chip-adb3115c) — [Stratechery](https://stratechery.com/2026/apple-updates-mini-and-studio-ai-computers-openai-jalapeno/)
- [Anthropic Adds Claude Mythos 5 to Claude Security for Vulnerability Scanning](https://www.wortins.com/story/anthropic-adds-claude-mythos-5-to-claude-security-for-vulner-34ff7302) — [Anthropic](https://claude.com/blog/bringing-claude-mythos-5-to-more-defenders)
- [Fundraisly](https://www.wortins.com/story/fundraisly-75fbe87d) — [Futurepedia](https://futurepedia.io/)
- [Alibaba Raises $10.2 Billion in Record Hong Kong Share Sale to Fund AI Expansion](https://www.wortins.com/story/alibaba-raises-10-2-billion-in-record-hong-kong-share-sale-t-ee1afa0a) — [Bloomberg](https://www.bloomberg.com/news/articles/2026-08-23/alibaba-to-raise-10-billion-by-selling-shares-for-ai-expansion)

---

_Curated and written by [Wortins](https://www.wortins.com) — The daily AI briefing. Every story links to its original source; the "Wortins read" on each is our own original analysis. [About Wortins & our editorial approach](https://www.wortins.com/about)._
