5 Sources
[1]
DeepSeek raises some V4 prices by more than 10x as AI demand strains capacity
DeepSeek is raising its prices, to the dismay of developers, but off-peak pricing, cache discounts, and multi-model routing mean the impact is more nuanced. One of AI vendor DeepSeek's biggest selling points has been its ultra-low price point, but that party's about to end. The Chinese model
[2]
DeepSeek's AI models are about to cost four times more - Engadget
DeepSeek's AI models are about to cost four times more Prices at off-peak hours will be half the peak-hour pricing. DeepSeek made its name offering far cheaper AI services than its pricier western competitors, but the cheap ride appears to be at an end. The company has started telling customers
[3]
DeepSeek officially launches V4-Pro AI model in August 2026
The Chinese AI startup released V4-Pro on its app, web, and API on Thursday, with a price increase set to follow on August 16 DeepSeek officially released its V4-Pro model on Thursday, making it available across the company's app, web interface, and API after the model had been in preview since
[4]
DeepSeek V4 Pro Launches at $0.435 per Million Input Tokens
DeepSeek V4 Pro has officially launched and its arrival is already making waves across the AI industry. Priced at just $0.435 per million input tokens and $0.87 per million output tokens, it sets a new standard for affordability without sacrificing performance. According to Universe of AI, the
[5]
DeepSeek rolls out V4-Pro with improved agentic AI, DSpark decoding and Codex support
DeepSeek has launched DeepSeek-V4-Pro-0813, the official release of DeepSeek-V4-Pro that supersedes the preview version. The model is built on the DeepSeek-V4-Pro Preview architecture with a DSpark speculative decoding module and includes improvements for agentic and production workloads. DeepSeek
Share
Copy Link
DeepSeek officially released its V4 Pro AI model with significant price increases taking effect August 16. The V4 Pro will cost $3.96 per million output tokens at peak hours, up from $0.87, while introducing half-price off-peak rates. Despite the fourfold increase, DeepSeek remains cheaper than competitors like OpenAI's GPT-5.6 Sol at $30 per million tokens.
DeepSeek has officially launched DeepSeek-V4-Pro-0813, ending the promotional pricing that made the Chinese AI startup a disruptive force in the market
1
. The general availability version supersedes the preview model released in April and introduces substantial changes to the company's pricing structure3
. Starting at 16:00 UTC on August 16, API pricing for the V4 model family will increase by notable margins, with some rates rising more than 1,100%1
. The V4 Pro will cost $3.96 for 1 million output tokens at peak hours, more than four times the current rate of $0.872
. Input tokens are priced at $0.435 per million4
. DeepSeek is introducing peak and off-peak rates to allocate resources more reasonably as AI demand strains infrastructure capacity2
.The new pricing structure includes off-peak rates at half the peak-hour price, encouraging developers toward more flexible workload scheduling
1
. During off-peak hours, V4 Pro output tokens will cost $1.98 per million, while the V4-Flash model will be available for $0.66 per million output tokens compared to $1.32 during peak hours2
. The current API pricing was originally supposed to be a promotion ending on May 31, but DeepSeek announced that month it would make the discounted prices permanent before ultimately deciding to proceed with the price hikes2
. DeepSeek had offered a 75% promotional discount on V4 Pro through May 5, making the pricing shift a notable reversal of the company's earlier strategy3
.
Source: Geeky Gadgets
Despite the fourfold increase, DeepSeek remains significantly cheaper than many competitors in the AI model pricing landscape
2
. Kimi K3, developed by Chinese company Moonshot, costs $15 per 1 million output tokens, while Anthropic's Fable 5 charges $50 per million output tokens3
. OpenAI's most advanced model, GPT-5.6 Sol, costs $30 for 1 million output tokens2
. DeepSeek V4 Pro is up to 57 times cheaper than Fable 5 while delivering nearly identical results in key benchmark scores4
. OpenAI's low-cost offering, GPT-5.6 Luna, costs $1.20, which is less than the DeepSeek V4-Flash at peak hours but more expensive than the off-peak rate2
.The V4-Pro-0813 focuses on agent capabilities, where AI systems use tools, execute code, and complete multi-step workflows without human intervention
3
. Benchmark scores demonstrate substantial improvements over the preview version across agentic AI performance metrics5
. The model scored 87.9 on Terminal Bench 2.1 compared to 72.1 for V4 Pro Preview, while achieving 62.7 on DeepSWE and 61.5 on NL2Repo3
5
. On the CyberGym benchmark, it achieved 83.3, narrowly surpassing Fable 5's score of 83.14
. The model also scored 74.1 on Toolathlon-Verified, 25.7 on Agents' Last Exam, and 31.8 on AutomationBench, outperforming Fable 5's 29.14
5
.
Source: Engadget
The V4 Pro API has been updated to work with the OpenAI Responses API format out of the box, providing native compatibility for developers
3
. Built-in Codex support allows for one-click setup, optimizing the model for coding workflows5
. Thinking effort levels for both V4 Pro and V4-Flash have been expanded to three settings: low for simple tasks, high for daily agent workflows, and max for complex tasks requiring more deliberation3
5
. The model is capable of handling a context window of up to 1 million tokens and can produce outputs as long as 384,000 tokens, with the option to run in either thinking or non-thinking mode3
.Related Stories
DeepSeek-V4-Pro-0813 includes DSpark speculative decoding, a module built on the V4 Pro Preview architecture that improves inference efficiency
5
. Developers can enable DSpark with vLLM using seven speculative tokens with greedy draft sampling, or through SGLang without specifying a separate draft model path since target and draft weights come from the same checkpoint5
. For local deployment, DeepSeek recommends temperature settings of 1.0 with top-p at 0.95 for agentic scenarios and top-p at 1.0 for other scenarios, with maximum output length of 384K tokens for high and max reasoning modes5
.
Source: InfoWorld
DeepSeek's R1 model went viral in early 2025 and briefly gave the company a commanding position in the AI race, but competitors including Moonshot AI, Alibaba, and ByteDance have since closed the gap with their own releases
3
. The launch comes after the cheaper V4-Flash model performed above expectations in independent tests, in some cases outpacing the April preview of V4 Pro, drawing attention given that Pro is positioned as the more capable product3
. Models like SpaceX's Grock 4.6, priced at $2 per million input tokens and $6 per million output tokens, offer competitive performance and match GPT-5.6 Sol with a score of 61 on the Artificial Analysis Index4
. The company secured roughly $7.4 billion in its first round of outside capital in June, and was reportedly in discussions for an additional round that would value it at around $74 billion3
. The model weights are released under the MIT License, maintaining accessibility for developers5
.Summarized by
Navi
[4]
06 Aug 2026•Business and Economy

06 Aug 2026•Business and Economy

24 Apr 2026•Technology

1
Policy and Regulation

2
Technology

3
Technology
