# AWS cuts GPT-5.6 Luna pricing by 80% in Bedrock, down to $0.20 per million tokens

> Amazon's Bedrock has slashed the price of running OpenAI's GPT-5.6 Luna models by up to 80 percent, dropping input tokens to $0.20 per million and output tokens to $1.20 per million. For anyone running high-volume workloads, that is the difference between a prototype and a product, and it continues the steady collapse in the cost of frontier-class inference. The cut lands in a busy pricing week. AWS also flagged Claude Sonnet 5 launching at $2 and $10 per million tokens through the end of August, though with a catch worth reading twice: Sonnet 5's new tokenizer generates about 30 percent more billable tokens, quietly clawing back part of any headline discount. The takeaway for builders is that sticker prices are getting harder to compare. Raw per-token rates are falling fast, but tokenizer changes, output-heavy pricing, and model routing all shape the real bill. Cheaper inference is unlocking use cases that were uneconomical a year ago, and the vendors know it.

_Section: [Daily AI Updates](https://www.wortins.com/daily-ai) · Source: AWS · Published Sunday, August 16, 2026_

## Wortins' read

Amazon's Bedrock has slashed the price of running OpenAI's GPT-5.6 Luna models by up to 80 percent, dropping input tokens to $0.20 per million and output tokens to $1.20 per million. For anyone running high-volume workloads, that is the difference between a prototype and a product, and it continues the steady collapse in the cost of frontier-class inference. The cut lands in a busy pricing week. AWS also flagged Claude Sonnet 5 launching at $2 and $10 per million tokens through the end of August, though with a catch worth reading twice: Sonnet 5's new tokenizer generates about 30 percent more billable tokens, quietly clawing back part of any headline discount. The takeaway for builders is that sticker prices are getting harder to compare. Raw per-token rates are falling fast, but tokenizer changes, output-heavy pricing, and model routing all shape the real bill. Cheaper inference is unlocking use cases that were uneconomical a year ago, and the vendors know it.

## Source

[Read the full story at AWS](https://aws.amazon.com/blogs/aws/aws-weekly-roundup-price-reduction-of-gpt-models-in-bedrock-cloudwatch-managed-collectors-for-prometheus-metrics-and-more-august-3-2026/)

## Related coverage

- [Emerald AI Raises $150 Million Series A at $1.05 Billion Valuation](https://www.wortins.com/story/emerald-ai-raises-150-million-series-a-at-1-05-billion-valua-2071899e) — [Business Wire](https://www.businesswire.com/news/home/20260825127649/en/Emerald-AI-Raises-$150-Million-Series-A-at-$1.05-Billion-Valuation-to-Scale-Power-Flexible-AI-Data-Centers)
- [ShepHertz Launches AgentAnywhere, Sovereign Agentic AI Platform for Regulated Industries](https://www.wortins.com/story/shephertz-launches-agentanywhere-sovereign-agentic-ai-platfo-a80b8050) — [ShepHertz](https://aiagentstore.ai/ai-agent-news/this-week)
- [Apple Updates, AI Computers, and OpenAI's Jalapeño Chip](https://www.wortins.com/story/apple-updates-ai-computers-and-openai-s-jalape-o-chip-adb3115c) — [Stratechery](https://stratechery.com/2026/apple-updates-mini-and-studio-ai-computers-openai-jalapeno/)
- [Anthropic Adds Claude Mythos 5 to Claude Security for Vulnerability Scanning](https://www.wortins.com/story/anthropic-adds-claude-mythos-5-to-claude-security-for-vulner-34ff7302) — [Anthropic](https://claude.com/blog/bringing-claude-mythos-5-to-more-defenders)
- [Fundraisly](https://www.wortins.com/story/fundraisly-75fbe87d) — [Futurepedia](https://futurepedia.io/)
- [Alibaba Raises $10.2 Billion in Record Hong Kong Share Sale to Fund AI Expansion](https://www.wortins.com/story/alibaba-raises-10-2-billion-in-record-hong-kong-share-sale-t-ee1afa0a) — [Bloomberg](https://www.bloomberg.com/news/articles/2026-08-23/alibaba-to-raise-10-billion-by-selling-shares-for-ai-expansion)

---

_Curated and written by [Wortins](https://www.wortins.com) — The daily AI briefing. Every story links to its original source; the "Wortins read" on each is our own original analysis. [About Wortins & our editorial approach](https://www.wortins.com/about)._
