DeepSeek raises V4 model prices, adding dynamic pricing from August 16; V4-Flash output tokens go from $0.28/1M to $1.32 during peak hours and $0.66 in off-peak
DeepSeek is steeply raising the prices for its flagship V4 models ahead of a potential initial public offering …
Bloomberg Saritha Rai
Context & Ripple Effects
DeepSeek’s V4 pricing has moved sharply from its earlier low-cost positioning: V4-Flash was listed at $0.28 per million output tokens in April, while V4-Pro received a permanent 75% API-price reduction in May. The company then signaled broad AI-service increases this month before setting out the peak and off-peak schedule.
The change arrives alongside the launch of V4-Pro and reporting that DeepSeek is considering an IPO, making monetization and capacity management more consequential than maintaining a single headline-low API price.
First-order effects
- DeepSeek API customers using V4-Flash face output-token costs of $1.32 per million at peak times and $0.66 off-peak, replacing the prior $0.28 rate and creating an immediate incentive to move deferrable workloads outside peak periods.
- DeepSeek gains a mechanism to charge more when demand is highest, rather than applying one rate across all V4-Flash usage.
Second-order effects
- API buyers will need to reassess V4-Flash budgets by workload timing, not just token volume, and may reserve peak-hour use for tasks whose value supports the higher output price.
- DeepSeek’s V4-Pro launch now sits beside a more segmented V4-Flash tariff, pushing customers to compare model choice and execution time together rather than treating the cheaper model as a fixed-cost default.
Third-order effects
- If other AI API providers follow this approach, capacity-aware pricing would make inference procurement resemble a scheduling problem, with customers optimizing when workloads run as well as which model they use.
- The reversal from May’s permanent V4-Pro discount to August price increases points to AI providers testing whether low introductory pricing can give way to pricing that better reflects constrained serving capacity and margins.
The trend: AI model providers are shifting from uniformly low API rates toward capacity-aware pricing that monetizes peak inference demand.
Related: Capacity-aware pricing · AI procurement phase · DeepSeek · DeepSeek’s planned AI-service price increases · DeepSeek’s V4-Pro price reduction
Related Coverage
- Change Log — The GA release of DeepSeek-V4-Pro has been rolled out on the APP, Web, and API. DeepSeek API Docs
- DeepSeek Announces Price Hikes for V4 Models The Information · Juro Osawa
- OpenAI and Anthropic in price war as Chinese AI rivals gain ground Australian Financial Review · Jamie John
- DeepSeek Lifts AI Model Prices Fourfold Wall Street Journal · Tracy Qu
- DeepSeek raises API pricing for its V4 models Reuters
- DeepSeek Harness — DeepSeek Harness (dsh) is an open-source agent harness developed by DeepSeek AI. GitHub
- DeepSeek-V4-Pro-0813 — DeepSeek-V4-Pro-0813 is the official release of DeepSeek-V4-Pro … Hugging Face
- Silicon Data: Closed US Models Lose Token Share to Chinese Open Models Blockchain.News
- DeepSeek's updated V4 Pro AI model struggles on benchmarks, shines in cybersecurity South China Morning Post · Xinmei Shen
- DeepSeek V4 Pro 0813 deepseek/deepseek-v4-pro-0813 DeepSeek on OpenRouter
- DeepSeek V4 Pro 0813 Hacker News
- DeepSeek-V4-Pro GA Release DeepSeek API Docs
- DeepSeek Raises API Prices Up to 12-Fold as China's AI Firms Chase Profit Seoul Economic Daily · Da-eun Jeong
- DeepSeek raises some V4 prices by more than 10x as AI demand strains capacity CIO.com · Taryn Plumb
- DeepSeek peak/off-peak pricing update Hacker News
- DeepSeek's AI models are about to cost four times more Engadget · Mariella Moon
- DeepSeek launches V4-Pro, its most advanced model that rivals Kimi K3 on some benchmarks but is priced much lower, at $0.435/1M input and $0.87/1M output tokens The Information
- OpenAI and Anthropic Slash AI Prices as Chinese Competitors Disrupt the Market Blockonomi · Trader Edge
Analysis
Discussion
-
@deepseek_ai
@deepseek_ai
on x
API pricing update 💰 With the V4 lineup release, we're updating our API pricing and introducing peak and off-peak rates. Off-peak rates are 50% lower than peak, enabling more flexible workload scheduling. 📉 New pricing takes effect at 16:00 UTC, Aug 16, 2026 🕒 [image]
-
@dodoreach
@dodoreach
on x
so for $20 now you can have: DeepSeek API: - now anywhere from 10M to 2.8 BILLION tokens in off-peak rates - you decide when your limit expires - no “5-hour limits” ever - nobody else's reset schedule dictates when you work - you know what you pay for - you know how much usage
-
@jun_song
Jun Song
on x
Official announcement from Deepseek. Also API price got updated. $0.66/$1.98 input/output Testing it right now if it's a different model from yesterday.
-
@teortaxestex
@teortaxestex
on x
in/out are almost irrelevant in the face of 2.5-12.2x increase in cache hit prices. This will obliterate their main current advantage. look, already $0,34 of cache hits (peak) alone in this quick session. They need to make DSH very powerful indeed. I think that's the entire plan.…
-
@silicon_data
@silicon_data
on x
DeepSeek raises model prices 4 times. As we had observed last week, an interesting trend in AI model layer is that open models are becoming more expensive while closed frontier models are becoming cheaper with successive price changes from ChatGPT, Grok and Muse Spark not to [ima…
-
@pigeon__s
@pigeon__s
on x
DeepSeek of all people more expensive than LUNA?!
-
@markgurman
Mark Gurman
on x
@Reuters This has been known and public for two years ...
-
@ns123abc
Nik
on x
🚨 DeepSeek new API prices just dropped >~2x the price hike on OFF-PEAK hours >peak hours are 2x on top of that >3-4X higher than what you pay right now poors in shambles [image]
-
@cheatyyyy
@cheatyyyy
on x
DeepSeek V4 Pricing has been 2x'd for off-peak and 4x'd for Peak hours on the official API [image]
-
@deepseek_ai
@deepseek_ai
on x
We're launching DeepSeek-V4-Pro today! 🚀 🔷 Major Agent upgrades with strong production gains! 🔷 Flexible reasoning effort for V4-Pro & V4-Flash: low for simple tasks, high for daily Agent workflows, max for complex tasks. 🔷 Native OpenAI Responses API support, optimized for Co…
-
@deepseek_ai
@deepseek_ai
on x
🧩 DeepSeek Harness v0.1 is now available in Developer Preview! 🔹 We're opening it up to developers building agent harnesses worldwide and open-sourcing the codebase in MIT license. 🔹 Powered by the Cordis meta-framework, DeepSeek Harness is an agent harness built around one
-
@altryne
Alex Volkov
on x
New DeepSeek V4 0813 dropped on @huggingface , my @bot sent an alert to our slack team, and the team said “he hallucinated the link” But @bot was good! Apparently @deepseek_ai published the weights (with MIT license!) and then took them down!? [image]
-
@eliebakouch
Elie
on x
deepseek harness was heavily developed using codex, at least ~20% of commits and PRs are coming from codex worktrees [image]
-
@scaling01
@scaling01
on x
DeepSeek is washed but their distillation seems good
-
@eliebakouch
Elie
on x
new deepseek v4 pro is now open weight on hugging face (mit license) “v4” is a bit misleading, previous model was only a preview and this one has way more training behind it, feels more like a v4.5 also very excited for the deepseek harness release https://huggingface.co/... [ima…
-
@cgtwts
@cgtwts
on x
“Sir... a new DeepSeek model just dropped. It reportedly delivers Fable 5-level performance at 57x cheaper cost”. [image]
-
@scaling01
@scaling01
on x
DeepSeek has been very disappointing this year V4 came much later than expected and was smaller than expected. The price hikes also hurt. They are currently 4th or 5th place in China
-
@scaling01
@scaling01
on x
I still don't see any weights for DeepSeek V4 and the pricing update is really bad [image]
-
@kimmonismus
@kimmonismus
on x
DeepSeek-V4-Pro GA officially announced! Yesterday's leaked benchmarks were real: The model now supports adjustable reasoning effort: low for simple tasks, high for everyday agent work, and max for complex problems. V4-Flash gets the same control. DeepSeek also added native [imag…
-
@thegenioo
Hamza
on x
DeepSeek has made the official announcement for V4 Pro (already out since yesterday) They have also made the pricing change (higher now) [image]
-
@aibattle_
@aibattle_
on x
New DeepSeek peak / off-peak pricing is live on the docs page. The new pricing will take effect on August 16 [image]
-
@sheriyuo
Xiuyu Li
on x
So AA 53 for V4 Pro 0813 may be inaccurate. 55-56 seems more convincing 👀 [image]
-
@ns123abc
Nik
on x
reminder that deepseek made NO announcement for the new V4 Pro anywhere on social media because they are ashamed of it >scores only 1 point higher than V4 Flash [image]
-
@arena
@arena
on x
DeepSeek-V4-Pro (Max) by @deepseek_ai is expected to shift the Pareto curve for Code Arena: WebDev with this upcoming open weights model. It currently sits at ~#8 overall (AutoEval) at 1607 pts. Priced at $0.435 input/ $0.87 output per MTokens, it outperforms models beyond its [i…
-
@teortaxestex
@teortaxestex
on x
Really weird things going on with V4-Pro-0813 release I will withhold judgement until they admit it's rolled out and post their own evals in English. Currently, the WeChat leak does not match AA even directionally (-2.7% for Flash, -9% for Pro). I doubt this is their best. [image…
-
@cheatyyyy
@cheatyyyy
on x
it's not a bad model, it is unfortunately definitely behind the frontier by a solid margin, but it feels a lot better than v4 pro preview to use, definitely past the opus 4.6 to opus 4.7 barrier for long tasks, much more reliable and the model has a good intuition. gpt-shaped.
-
@teortaxestex
@teortaxestex
on x
V4-Pro-0813 massively surpasses 0731 on goonbench (internal) one niche already secured
-
@goodhunt
Hunter Bown
on x
for some reason the chinese internet hates the new ds v4 pro I don't get it
-
@lentils80
@lentils80
on x
Deepseek V4 Pro GA scoring only 1 point higher than Flash on Artificial Analysis while being like 5 times larger is crazy work [image]
-
@zephyr_z9
@zephyr_z9
on x
IMO, they have finally stabilized the new V4 arch and are probably preparing for a larger 3T-4T V4 Max or V5 I don't think it makes a lot of sense to spend resources on V4 Pro post training
-
@scaling01
@scaling01
on x
today is a good day: DeepSeek-V4-Pro Grok 4.6 Qwen3.8-2.4T and maybe something from OpenAI
-
@hosseeb
Haseeb
on x
Wow. @deepseek_ai 's V4-Pro just launched via API and it looks incredible. This thing costs $0.87/M out. Opus is $25/M out. It's almost 30x cheaper than Opus! Benchmarks looks like it's somewhere between Opus 4.8 and Opus 5 quality. Absolutely incredible deflation in frontie…
-
@openrouter
@openrouter
on x
DeepSeek V4 Pro 0813 is live on OpenRouter. @deepseek_ai reports large agent gains over V4 Pro Preview: DeepSWE 62.7 (+49.9), CyberGym 83.3 (+30.6), NL2Repo 61.5 (+23.0), and Terminal Bench 2.1 87.9 (+15.8) More providers coming online soon Use it now: https://openrouter.ai/...
-
@sakurayukiai
Sakura Yuki
on x
1M context and a 384K max output turn DeepSeek V4 Pro 0813 into a scheduler puzzle. Does the provider reserve KV capacity against the full cap, or allocate blocks lazily and risk eviction when several generations refuse to stop?
-
@andrewcurran_
Andrew Curran
on x
As yet unverified DeepSeek V4 Pro benchmarks from WeChat. Seismic if accurate. [image]
-
@arena
@arena
on x
Big news: DeepSeek-V4-Pro (Max) by @deepseek_ai is coming in around ~#8 overall (#2 among open models) in the Code Arena: WebDev! At 1607 pts, this places it after GPT-5.6 Sol (xHigh)(1622 pts), and makes it the second best open model after Kimi K3 (Max) (1674 pts). Note: this [i…
-
r/LocalLLaMA
r
on reddit
Deepseek Harness is Up!
-
r/LocalLLaMA
r
on reddit
GitHub - deepseek-ai/deepseek-harness
-
r/DeepSeek
r
on reddit
DeepSeek Harness is out !