DeepSeek Raises V4 API Prices by Up to 1,100% as AI Demand Strains Capacity
By Vikram Singh
Updated on Aug 14, 2026 | 5 min read | 1.43K+ views
Share:
All courses
Certifications
More
By Vikram Singh
Updated on Aug 14, 2026 | 5 min read | 1.43K+ views
Share:
Table of Contents
Key Highlights of NEWS
AI models are becoming more powerful, but their cost and infrastructure needs are also evolving. Explore our Artificial Intelligence course to understand modern AI models, applications, and the technologies driving this shift.
Chinese AI company DeepSeek is raising API prices for its V4-Pro and V4-Flash models.
The new rates will take effect on August 17. They will range from 50% to 1,100% above current prices. The exact increase depends on the model, token type, and usage period.
The company is also introducing peak and off-peak pricing. The change marks a shift for DeepSeek. The company has built much of its reputation around offering powerful AI models at relatively low prices.
Popular AI Programs
DeepSeek will use different rates for peak and off-peak periods. The company will continue to charge separately for input tokens, output tokens, and cached input tokens.
Under the new structure, peak-period usage will cost more than off-peak usage. The increases vary significantly across the two V4 models.
Model |
Token type |
Current price |
New off-peak price |
New peak price |
| V4-Flash | Cached input | $0.0028 | $0.007 | $0.014 |
| V4-Flash | Input | $0.14 | $0.22 | $0.44 |
| V4-Flash | Output | $0.28 | $0.66 | $1.32 |
| V4-Pro | Cached input | $0.003625 | $0.022 | $0.044 |
| V4-Pro | Input | $0.435 | $0.66 | $1.32 |
| V4-Pro | Output | $0.87 | $1.98 | $3.96 |
The figures show why some reports describe the change as a more than 10x increase.
The largest increase applies to cached V4-Pro input during peak periods. The price rises from $0.003625 to $0.044 per million tokens. That represents an increase of more than 1,100%.
DeepSeek's current API documentation lists V4-Pro at $0.003625 per million cached input tokens, $0.435 for uncached input, and $0.87 for output. V4-Flash is priced at $0.0028, $0.14, and $0.28 respectively.
The price increase comes as demand for AI services continues to rise. DeepSeek's V4 models have attracted developers because of their performance and low API costs. The company's V4 family supports a 1-million-token context window. Both V4-Pro and V4-Flash also support thinking and non-thinking modes.
DeepSeek has also positioned V4 around agentic AI and coding workloads. V4-Pro is designed for complex reasoning and agent tasks. V4-Flash targets faster and more cost-efficient workloads.
Higher demand can increase the cost of running these models. AI inference requires large amounts of computing capacity, particularly for workloads involving long contexts and repeated model calls.
The new pricing structure gives DeepSeek another way to manage that demand. It also encourages developers to move less urgent workloads to cheaper off-peak periods.
As AI models become more complex and demand for AI services grows, machine learning expertise is becoming increasingly valuable. The Master of Science in Machine Learning & AI from LJMU can help you build advanced skills in machine learning, AI, and real-world AI applications.
AI Courses to upskill
Explore Artificial Intelligence Courses for Career Progression
The price increase could raise operating costs for developers using DeepSeek's API at scale. The impact will depend on how applications use the models. Developers with high volumes of cached input tokens could see a particularly large change in costs. V4-Pro users could also face higher output costs under the new pricing structure.
The peak and off-peak model may encourage companies to change when they run AI workloads. For example, batch processing and other non-urgent tasks could be shifted to off-peak periods. The change could also affect companies comparing DeepSeek with other AI providers.
DeepSeek has historically competed on AI performance and low inference costs. Higher API prices could reduce part of that cost advantage.
However, the company is still offering lower rates during off-peak periods. The actual cost will depend on usage patterns, token consumption, and the model selected.
DeepSeek's move also highlights a broader issue in the AI industry. As AI agents become more capable, they can make many API calls during a single task. This can increase inference demand quickly.
Pricing is therefore becoming an important factor in deciding which AI models developers use in production.
DeepSeek's V4 API price increase marks a major change in its pricing strategy. The company is moving from a simple low-cost model to peak and off-peak pricing. The increases can reach 1,100% for some token categories.
The move also reflects the rising cost of serving high-demand AI workloads. For developers, API pricing will become an even more important factor when choosing models for production AI applications and agentic workflows.
DeepSeek is raising prices for V4-Pro and V4-Flash as demand for its AI services increases. The company is also introducing peak and off-peak pricing.
The new pricing will take effect on August 17, 2026, according to Reuters' report on DeepSeek's announcement.
The increases range from 50% to 1,100%, depending on the model, token type, and usage period.
The largest increase applies to V4-Pro cached input during peak periods. The price rises from $0.003625 to $0.044 per million tokens.
Under the new structure, V4-Pro costs $0.66 per million uncached input tokens and $1.98 per million output tokens during off-peak periods. Peak pricing rises to $1.32 for input and $3.96 for output.
V4-Flash will cost $0.22 per million uncached input tokens and $0.66 per million output tokens during off-peak periods. Peak rates will be $0.44 for input and $1.32 for output.
Peak and off-peak pricing charges different rates based on when an API request is processed. Peak periods have higher prices, while off-peak periods offer lower rates.
DeepSeek V4 offers a large context window, reasoning capabilities, tool calling, and support for agentic workloads. It has also been positioned as a cost-efficient alternative to other advanced AI models.
Yes. Companies using DeepSeek APIs at high volumes may see higher operating costs. The impact will depend on their token usage, caching patterns, model choice, and usage time.
DeepSeek's prices remain competitive, especially during off-peak periods. However, the increase reduces the cost advantage that helped make its V4 models attractive to developers.
131 articles published
Vikram Singh is a seasoned content strategist with over 5 years of experience in simplifying complex technical subjects. Holding a postgraduate degree in Applied Mathematics, he specializes in creatin...
Speak with AI & ML expert
By submitting, I accept the T&C and
Privacy Policy
Top Resources