# OpenAI Introduces Ultrafast Mode: GPT-5.6 Sol Now Runs 14x Faster

> OpenAI has launched Ultrafast, a new mode that runs its GPT-5.6 Sol model at up to 750 output tokens per second, roughly 14 times its standard speed. The gain does not come from a smaller or weaker model but from where it runs. OpenAI is serving Ultrafast on hardware from Cerebras, whose wafer-scale chips keep an entire model on a single giant piece of silicon and avoid much of the memory shuttling that slows conventional GPUs. Speed changes what a model is useful for. At conversational latency a lot of waiting disappears, which matters most for jobs where a person or another system is blocked until the answer arrives. OpenAI is aiming Ultrafast at exactly those cases, citing incident response, customer service, financial analysis, and e-commerce. For now it is a limited preview for enterprise customers, with access widening as capacity grows, and Anthropic offers a comparable fast mode for Claude at somewhat lower speeds. The broader signal is that inference speed, not just raw intelligence, is becoming a competitive axis, and specialized chips are how the labs intend to win it.

_Section: [Daily AI Updates](https://www.wortins.com/daily-ai) · Source: TechCrunch · Published Saturday, August 15, 2026_

## Wortins' read

OpenAI has launched Ultrafast, a new mode that runs its GPT-5.6 Sol model at up to 750 output tokens per second, roughly 14 times its standard speed. The gain does not come from a smaller or weaker model but from where it runs. OpenAI is serving Ultrafast on hardware from Cerebras, whose wafer-scale chips keep an entire model on a single giant piece of silicon and avoid much of the memory shuttling that slows conventional GPUs. Speed changes what a model is useful for. At conversational latency a lot of waiting disappears, which matters most for jobs where a person or another system is blocked until the answer arrives. OpenAI is aiming Ultrafast at exactly those cases, citing incident response, customer service, financial analysis, and e-commerce. For now it is a limited preview for enterprise customers, with access widening as capacity grows, and Anthropic offers a comparable fast mode for Claude at somewhat lower speeds. The broader signal is that inference speed, not just raw intelligence, is becoming a competitive axis, and specialized chips are how the labs intend to win it.

## Source

[Read the full story at TechCrunch](https://techcrunch.com/2026/08/13/openai-introduces-ultrafast-a-new-mode-that-makes-gpt-5-6-sol-work-at-14x-the-speed/)

## Related coverage

- [Emerald AI Raises $150 Million Series A at $1.05 Billion Valuation](https://www.wortins.com/story/emerald-ai-raises-150-million-series-a-at-1-05-billion-valua-2071899e) — [Business Wire](https://www.businesswire.com/news/home/20260825127649/en/Emerald-AI-Raises-$150-Million-Series-A-at-$1.05-Billion-Valuation-to-Scale-Power-Flexible-AI-Data-Centers)
- [Agents Over Bubbles](https://www.wortins.com/story/agents-over-bubbles-3e7bd726) — [Stratechery](https://stratechery.com/2026/agents-over-bubbles/)
- [Anthropic Adds Claude Mythos 5 to Claude Security for Vulnerability Scanning](https://www.wortins.com/story/anthropic-adds-claude-mythos-5-to-claude-security-for-vulner-34ff7302) — [Anthropic](https://claude.com/blog/bringing-claude-mythos-5-to-more-defenders)
- [OpenAI Expands Daybreak With GPT-5.6-Cyber Cybersecurity Model](https://www.wortins.com/story/openai-expands-daybreak-with-gpt-5-6-cyber-cybersecurity-mod-9f70603c) — [TechCrunch](https://techcrunch.com/2026/08/10/as-ai-led-attacks-multiply-openai-launches-a-new-cyber-model/)
- [Fundraisly](https://www.wortins.com/story/fundraisly-75fbe87d) — [Futurepedia](https://futurepedia.io/)
- [ShepHertz Launches AgentAnywhere, Sovereign Agentic AI Platform for Regulated Industries](https://www.wortins.com/story/shephertz-launches-agentanywhere-sovereign-agentic-ai-platfo-a80b8050) — [ShepHertz](https://aiagentstore.ai/ai-agent-news/this-week)

---

_Curated and written by [Wortins](https://www.wortins.com) — The daily AI briefing. Every story links to its original source; the "Wortins read" on each is our own original analysis. [About Wortins & our editorial approach](https://www.wortins.com/about)._
