DeepSeek says it will lower V4 Pro API prices by 75% to $0.435/1M input and $0.87/1M output tokens, making permanent the discount prices set to expire on May 31
DeepSeek said it will make permanent a steep discount on its flagship V4-Pro model, maintaining prices for developers at a quarter of their original level.
Context & Ripple Effects
DeepSeek introduced V4 Pro and V4 Flash in preview in April, positioning V4 Pro several months behind the leading edge on performance while pricing both models aggressively. The new permanent V4 Pro rate extends a temporary discount rather than treating low pricing as a short-term launch promotion.
Earlier coverage showed DeepSeek using time-based API discounts and reporting high theoretical inference margins for prior models. That history makes the decision relevant as a sustained commercial positioning choice, even as the company is reportedly pursuing new capital, infrastructure expansion and an eventual China IPO.
First-order effects
- Developers using V4 Pro get a durable reduction in input and output inference costs, improving the economics of applications already built around the model.
- DeepSeek forgoes the scheduled reversion to its higher list price and anchors V4 Pro as a low-cost flagship offering alongside its cheaper V4 Flash model.
Second-order effects
- Rival API providers serving price-sensitive developers face a clearer benchmark: matching DeepSeek may require lower prices, differentiated capabilities, or commercial terms that reduce customers' switching incentives.
- Lower recurring token costs can make higher-volume or more margin-sensitive AI features more viable for DeepSeek customers, shifting demand toward inference capacity rather than one-off model evaluation.
Third-order effects
- If flagship-quality models continue to be offered at permanently compressed API rates, model access is likely to become less differentiated by price alone and more dependent on performance, reliability, tooling and distribution.
- The move underscores an inference-economics race: providers that can lower serving costs through infrastructure and hardware choices may have more room to sustain price pressure, though DeepSeek's long-term ability to do so will depend on execution and capital needs.
The trend: This is another data point in the shift from promotional AI API pricing toward sustained inference-price competition among model providers.