# DeepSeek V4 Model Launches Mid-July With Expanded Context and Peak-Hour Pricing

> DeepSeek is formally launching V4 in mid-July, extending the preview it showed in April into a full release. The model pushes its context window to one million tokens and claims stronger reasoning, agentic and coding performance, arriving in two flavors: deepseek-v4-flash for latency-sensitive work and deepseek-v4-pro for heavier reasoning. The more interesting wrinkle is the business model. DeepSeek is introducing peak-hour pricing, so calls made during busy local windows, roughly 9am to noon and 2pm to 6pm, cost twice the off-peak rate, while off-peak prices stay where they were. It is a surge-pricing approach borrowed from ride-hailing and cloud computing, and a candid acknowledgment that inference capacity is a scarce, time-sensitive resource. For a Chinese lab operating under US export controls, squeezing more value out of constrained hardware is not a gimmick, it is strategy. Peak-hour pricing nudges non-urgent workloads into off hours and smooths demand, which matters a great deal when you cannot simply buy your way to more chips. Expect other providers to watch closely.

_Section: [Daily AI Updates](https://www.wortins.com/daily-ai) · Source: Let's Data Science · Published Friday, July 10, 2026_

## Wortins' read

DeepSeek is formally launching V4 in mid-July, extending the preview it showed in April into a full release. The model pushes its context window to one million tokens and claims stronger reasoning, agentic and coding performance, arriving in two flavors: deepseek-v4-flash for latency-sensitive work and deepseek-v4-pro for heavier reasoning. The more interesting wrinkle is the business model. DeepSeek is introducing peak-hour pricing, so calls made during busy local windows, roughly 9am to noon and 2pm to 6pm, cost twice the off-peak rate, while off-peak prices stay where they were. It is a surge-pricing approach borrowed from ride-hailing and cloud computing, and a candid acknowledgment that inference capacity is a scarce, time-sensitive resource. For a Chinese lab operating under US export controls, squeezing more value out of constrained hardware is not a gimmick, it is strategy. Peak-hour pricing nudges non-urgent workloads into off hours and smooths demand, which matters a great deal when you cannot simply buy your way to more chips. Expect other providers to watch closely.

## Source

[Read the full story at Let's Data Science](https://letsdatascience.com/news/deepseek-plans-mid-july-v4-release-with-peak-hour-pricing-5187a98d)

## Related coverage

- [Ours Privacy Raises $15M Series A for Healthcare AI Data Platform](https://www.wortins.com/story/ours-privacy-raises-15m-series-a-for-healthcare-ai-data-plat-4ae00549) — [Crunchbase News](https://news.crunchbase.com/venture/biggest-funding-rounds-ai-defense-fintech-robotics/)
- [Rippling Launches AI Spend Console to Measure Employee Token ROI and Curtail Runaway Costs](https://www.wortins.com/story/rippling-launches-ai-spend-console-to-measure-employee-token-f2cda775) — [TechCrunch](https://techcrunch.com/2026/08/07/after-rippling-blew-millions-on-ai-in-months-it-built-an-employee-roi-tool/)
- [Rillet Reaches Unicorn Status with $100M Series C](https://www.wortins.com/story/rillet-reaches-unicorn-status-with-100m-series-c-58a69bce) — [TechCrunch](https://techcrunch.com/2026/08/21/how-ai-accounting-startup-rillet-raised-100m-and-became-a-unicorn-in-48-hours/)
- [GitHub Copilot Workspace Now Lets Developers Rewind Conversations and Fork Sessions](https://www.wortins.com/story/github-copilot-workspace-now-lets-developers-rewind-conversa-be6ac530) — [GitHub Blog](https://github.blog/changelog/2026-08-07-github-copilot-weekly-releases-august-3/)
- [Google Launches Gemini 3.7 Flash at Half the Previous Price](https://www.wortins.com/story/google-launches-gemini-3-7-flash-at-half-the-previous-price-9ff50c48) — [VentureBeat](https://venturebeat.com/technology/googles-gemini-3-7-flash-targets-coding-and-agents-with-a-50-introductory-price-cut/)
- [OpenAI Previews Ultrafast Mode for GPT-5.6 Sol: 14x Faster Processing Speed](https://www.wortins.com/story/openai-previews-ultrafast-mode-for-gpt-5-6-sol-14x-faster-pr-efceab00) — [TechCrunch](https://techcrunch.com/2026/08/13/openai-introduces-ultrafast-a-new-mode-that-makes-gpt-5-6-sol-work-at-14x-the-speed/)

---

_Curated and written by [Wortins](https://www.wortins.com) — The daily AI briefing. Every story links to its original source; the "Wortins read" on each is our own original analysis. [About Wortins & our editorial approach](https://www.wortins.com/about)._
