/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

DeepSeek raises V4 model prices, adding dynamic pricing from August 16; V4-Flash output tokens go from $0.28/1M to $1.32 during peak hours and $0.66 in off-peak

DeepSeek is steeply raising the prices for its flagship V4 models ahead of a potential initial public offering …

Bloomberg Saritha Rai

Context & Ripple Effects

DeepSeek’s V4 pricing has moved sharply from its earlier low-cost positioning: V4-Flash was listed at $0.28 per million output tokens in April, while V4-Pro received a permanent 75% API-price reduction in May. The company then signaled broad AI-service increases this month before setting out the peak and off-peak schedule.

The change arrives alongside the launch of V4-Pro and reporting that DeepSeek is considering an IPO, making monetization and capacity management more consequential than maintaining a single headline-low API price.

First-order effects

  • DeepSeek API customers using V4-Flash face output-token costs of $1.32 per million at peak times and $0.66 off-peak, replacing the prior $0.28 rate and creating an immediate incentive to move deferrable workloads outside peak periods.
  • DeepSeek gains a mechanism to charge more when demand is highest, rather than applying one rate across all V4-Flash usage.

Second-order effects

  • API buyers will need to reassess V4-Flash budgets by workload timing, not just token volume, and may reserve peak-hour use for tasks whose value supports the higher output price.
  • DeepSeek’s V4-Pro launch now sits beside a more segmented V4-Flash tariff, pushing customers to compare model choice and execution time together rather than treating the cheaper model as a fixed-cost default.

Third-order effects

  • If other AI API providers follow this approach, capacity-aware pricing would make inference procurement resemble a scheduling problem, with customers optimizing when workloads run as well as which model they use.
  • The reversal from May’s permanent V4-Pro discount to August price increases points to AI providers testing whether low introductory pricing can give way to pricing that better reflects constrained serving capacity and margins.

The trend: AI model providers are shifting from uniformly low API rates toward capacity-aware pricing that monetizes peak inference demand.

Discussion

  • @deepseek_ai @deepseek_ai on x
    API pricing update 💰 With the V4 lineup release, we're updating our API pricing and introducing peak and off-peak rates. Off-peak rates are 50% lower than peak, enabling more flexible workload scheduling. 📉 New pricing takes effect at 16:00 UTC, Aug 16, 2026 🕒 [image]
  • @dodoreach @dodoreach on x
    so for $20 now you can have: DeepSeek API: - now anywhere from 10M to 2.8 BILLION tokens in off-peak rates - you decide when your limit expires - no “5-hour limits” ever - nobody else's reset schedule dictates when you work - you know what you pay for - you know how much usage
  • @jun_song Jun Song on x
    Official announcement from Deepseek. Also API price got updated. $0.66/$1.98 input/output Testing it right now if it's a different model from yesterday.
  • @teortaxestex @teortaxestex on x
    in/out are almost irrelevant in the face of 2.5-12.2x increase in cache hit prices. This will obliterate their main current advantage. look, already $0,34 of cache hits (peak) alone in this quick session. They need to make DSH very powerful indeed. I think that's the entire plan.…
  • @silicon_data @silicon_data on x
    DeepSeek raises model prices 4 times. As we had observed last week, an interesting trend in AI model layer is that open models are becoming more expensive while closed frontier models are becoming cheaper with successive price changes from ChatGPT, Grok and Muse Spark not to [ima…
  • @pigeon__s @pigeon__s on x
    DeepSeek of all people more expensive than LUNA?!
  • @markgurman Mark Gurman on x
    @Reuters This has been known and public for two years ...
  • @ns123abc Nik on x
    🚨 DeepSeek new API prices just dropped >~2x the price hike on OFF-PEAK hours >peak hours are 2x on top of that >3-4X higher than what you pay right now poors in shambles [image]
  • @cheatyyyy @cheatyyyy on x
    DeepSeek V4 Pricing has been 2x'd for off-peak and 4x'd for Peak hours on the official API [image]
  • @deepseek_ai @deepseek_ai on x
    We're launching DeepSeek-V4-Pro today!  🚀 🔷 Major Agent upgrades with strong production gains!  🔷 Flexible reasoning effort for V4-Pro & V4-Flash: low for simple tasks, high for daily Agent workflows, max for complex tasks.  🔷 Native OpenAI Responses API support, optimized for Co…
  • @deepseek_ai @deepseek_ai on x
    🧩 DeepSeek Harness v0.1 is now available in Developer Preview! 🔹 We're opening it up to developers building agent harnesses worldwide and open-sourcing the codebase in MIT license. 🔹 Powered by the Cordis meta-framework, DeepSeek Harness is an agent harness built around one
  • @altryne Alex Volkov on x
    New DeepSeek V4 0813 dropped on @huggingface , my @bot sent an alert to our slack team, and the team said “he hallucinated the link” But @bot was good! Apparently @deepseek_ai published the weights (with MIT license!) and then took them down!? [image]
  • @eliebakouch Elie on x
    deepseek harness was heavily developed using codex, at least ~20% of commits and PRs are coming from codex worktrees [image]
  • @scaling01 @scaling01 on x
    DeepSeek is washed but their distillation seems good
  • @eliebakouch Elie on x
    new deepseek v4 pro is now open weight on hugging face (mit license) “v4” is a bit misleading, previous model was only a preview and this one has way more training behind it, feels more like a v4.5 also very excited for the deepseek harness release https://huggingface.co/... [ima…
  • @cgtwts @cgtwts on x
    “Sir... a new DeepSeek model just dropped. It reportedly delivers Fable 5-level performance at 57x cheaper cost”. [image]
  • @scaling01 @scaling01 on x
    DeepSeek has been very disappointing this year V4 came much later than expected and was smaller than expected. The price hikes also hurt. They are currently 4th or 5th place in China
  • @scaling01 @scaling01 on x
    I still don't see any weights for DeepSeek V4 and the pricing update is really bad [image]
  • @kimmonismus @kimmonismus on x
    DeepSeek-V4-Pro GA officially announced! Yesterday's leaked benchmarks were real: The model now supports adjustable reasoning effort: low for simple tasks, high for everyday agent work, and max for complex problems. V4-Flash gets the same control. DeepSeek also added native [imag…
  • @thegenioo Hamza on x
    DeepSeek has made the official announcement for V4 Pro (already out since yesterday) They have also made the pricing change (higher now) [image]
  • @aibattle_ @aibattle_ on x
    New DeepSeek peak / off-peak pricing is live on the docs page. The new pricing will take effect on August 16 [image]
  • @sheriyuo Xiuyu Li on x
    So AA 53 for V4 Pro 0813 may be inaccurate. 55-56 seems more convincing 👀 [image]
  • @ns123abc Nik on x
    reminder that deepseek made NO announcement for the new V4 Pro anywhere on social media because they are ashamed of it >scores only 1 point higher than V4 Flash [image]
  • @arena @arena on x
    DeepSeek-V4-Pro (Max) by @deepseek_ai is expected to shift the Pareto curve for Code Arena: WebDev with this upcoming open weights model. It currently sits at ~#8 overall (AutoEval) at 1607 pts. Priced at $0.435 input/ $0.87 output per MTokens, it outperforms models beyond its [i…
  • @teortaxestex @teortaxestex on x
    Really weird things going on with V4-Pro-0813 release I will withhold judgement until they admit it's rolled out and post their own evals in English. Currently, the WeChat leak does not match AA even directionally (-2.7% for Flash, -9% for Pro). I doubt this is their best. [image…
  • @cheatyyyy @cheatyyyy on x
    it's not a bad model, it is unfortunately definitely behind the frontier by a solid margin, but it feels a lot better than v4 pro preview to use, definitely past the opus 4.6 to opus 4.7 barrier for long tasks, much more reliable and the model has a good intuition. gpt-shaped.
  • @teortaxestex @teortaxestex on x
    V4-Pro-0813 massively surpasses 0731 on goonbench (internal) one niche already secured
  • @goodhunt Hunter Bown on x
    for some reason the chinese internet hates the new ds v4 pro I don't get it
  • @lentils80 @lentils80 on x
    Deepseek V4 Pro GA scoring only 1 point higher than Flash on Artificial Analysis while being like 5 times larger is crazy work [image]
  • @zephyr_z9 @zephyr_z9 on x
    IMO, they have finally stabilized the new V4 arch and are probably preparing for a larger 3T-4T V4 Max or V5 I don't think it makes a lot of sense to spend resources on V4 Pro post training
  • @scaling01 @scaling01 on x
    today is a good day: DeepSeek-V4-Pro Grok 4.6 Qwen3.8-2.4T and maybe something from OpenAI
  • @hosseeb Haseeb on x
    Wow. @deepseek_ai 's V4-Pro just launched via API and it looks incredible.  This thing costs $0.87/M out.  Opus is $25/M out.  It's almost 30x cheaper than Opus!  Benchmarks looks like it's somewhere between Opus 4.8 and Opus 5 quality.  Absolutely incredible deflation in frontie…
  • @openrouter @openrouter on x
    DeepSeek V4 Pro 0813 is live on OpenRouter. @deepseek_ai reports large agent gains over V4 Pro Preview: DeepSWE 62.7 (+49.9), CyberGym 83.3 (+30.6), NL2Repo 61.5 (+23.0), and Terminal Bench 2.1 87.9 (+15.8) More providers coming online soon Use it now: https://openrouter.ai/...
  • @sakurayukiai Sakura Yuki on x
    1M context and a 384K max output turn DeepSeek V4 Pro 0813 into a scheduler puzzle. Does the provider reserve KV capacity against the full cap, or allocate blocks lazily and risk eviction when several generations refuse to stop?
  • @andrewcurran_ Andrew Curran on x
    As yet unverified DeepSeek V4 Pro benchmarks from WeChat. Seismic if accurate. [image]
  • @arena @arena on x
    Big news: DeepSeek-V4-Pro (Max) by @deepseek_ai is coming in around ~#8 overall (#2 among open models) in the Code Arena: WebDev! At 1607 pts, this places it after GPT-5.6 Sol (xHigh)(1622 pts), and makes it the second best open model after Kimi K3 (Max) (1674 pts). Note: this [i…
  • r/LocalLLaMA r on reddit
    Deepseek Harness is Up!
  • r/LocalLLaMA r on reddit
    GitHub - deepseek-ai/deepseek-harness
  • r/DeepSeek r on reddit
    DeepSeek Harness is out !