DeepSeek Price Increase Warning: Resolved (August 2026)
DeepSeek has published its peak and off-peak API rate card. See what the August 6 warning meant and find the current pricing analysis.
By AI Pricing Guru Editorial Team
AI Pricing Guru articles are maintained by the editorial workflow behind the site: daily pricing snapshots, provider source checks, and review passes for model launches, subscription limits, and billing changes.
TL;DR
- Update: DeepSeek has now published the new rate card. Peak and off-peak billing starts at 16:00 UTC on August 16, 2026.
- The original August 6 warning is resolved; see our current DeepSeek peak-pricing analysis for the generated rate table and buyer advice.
- Today's official Flash and Pro prices remain active until the cutover; the live table below pulls them from our canonical dataset.
Current cost baseline before DeepSeek's announced increase
USD per 1M tokens. Input and output rates are charted separately.
Stress-test your monthly API budget
Assumes 75% input tokens and 25% output tokens using current per-million rates.
DeepSeek V4 Flash 0731
deepseek
$3.30
- Input share
- $1.65
- Output share
- $1.65
GPT-5.6 Luna
openai
$4.50
- Input share
- $1.50
- Output share
- $3.00
DeepSeek V4 Pro 0813
deepseek
$9.90
- Input share
- $4.95
- Output share
- $4.95
GLM-5.2
zai
$21.50
- Input share
- $10.50
- Output share
- $11.00
Current DeepSeek rates and practical API alternatives
| Model | Provider | Input / 1M | Cached / 1M | Output / 1M |
|---|---|---|---|---|
| DeepSeek V4 Flash 0731 | deepseek | $0.22 | $0.0070 | $0.66 |
| DeepSeek V4 Pro 0813 | deepseek | $0.66 | $0.022 | $1.98 |
| GLM-5.2 | zai | $1.40 | $0.26 | $4.40 |
| GPT-5.6 Luna | openai | $0.2 | $0.02 | $1.20 |
Built from pricing.json at publish time.
August 14 update: DeepSeek has now published the new rate card and effective time. Read our current DeepSeek peak and off-peak pricing analysis for the generated rate comparison, peak windows, and practical scheduling advice. The article below preserves what buyers knew when DeepSeek issued its original August 6 warning.
DeepSeek placed an unusually blunt warning on its official pricing documentation: it planned to raise overall API pricing in the near future, with a “significant increase expected.” At the time, the notice did not specify a percentage, new token rates, affected billing categories, or an effective date.
That distinction mattered at the time. A price increase was confirmed; its size and timing were not. The current rate card still applied on August 6, 2026, and AI Pricing Guru did not change its canonical DeepSeek pricing data to a speculative future rate.
What DeepSeek actually announced
| Pricing question | Confirmed status on August 6 |
|---|---|
| Is an increase planned? | Yes—DeepSeek says overall API pricing will rise |
| How large will it be? | Undisclosed; DeepSeek only says it expects a significant increase |
| When will it start? | Undisclosed |
| Which models or billing items change? | Undisclosed |
| Are today’s listed rates still active? | Yes—the official page continues to publish the current Flash and Pro rate card |
The warning appears as a note beneath the DeepSeek V4 pricing table. The provider also tells customers to plan usage accordingly and says the specific plan will be subject to a later official notice.
The claim reached Hacker News on August 6, but that discussion adds no verified rate card. Comments guessing which billing items will rise are analysis, not an announcement. In particular, buyers should not convert an unverified multiplier into a procurement forecast.
Pricing impact: uncertainty is now part of the bill
DeepSeek’s current advantage is not only cheap fresh input and output. Its automatic cache-hit rate makes stable, repeated prefixes especially inexpensive for agents, retrieval workflows, and long system prompts. A broad increase could therefore affect architectures built around both low output cost and aggressive cache reuse.
The live chart, table, and calculator above provide today’s baseline against DeepSeek V4 Flash, V4 Pro, Z.ai pricing, and OpenAI pricing. They do not predict DeepSeek’s next rate card. Run several scenarios in the token calculator, then compare total accepted-task cost rather than only the per-token line item.
Who benefits—and who loses
The clearest losers are teams that routed high-volume production traffic to DeepSeek on the assumption that current rates were durable. Coding agents can create large output volumes and repeated tool calls, so even a moderate increase can compound across retries and verification loops.
Budget owners also lose forecast confidence. DeepSeek has given advance warning, but without a date or rate it is impossible to lock a precise next-quarter API budget.
Competing providers and third-party hosts benefit. Open-weight DeepSeek models can be served elsewhere, while GLM, Kimi, Qwen, and budget OpenAI routes give buyers more negotiating and migration options. That does not guarantee a cheaper replacement: snapshot, context length, latency, cache behavior, and reliability must match the workload.
What DeepSeek API users should do now
- Export token usage by model, input type, cache status, and output volume; identify the workflows most exposed to a broad increase.
- Model several higher-cost scenarios without treating any one as DeepSeek’s announced price.
- Add a billing alert and recheck the official pricing page before each major top-up or contract decision.
- Canary one compatible fallback now, while migration is optional rather than urgent.
- Avoid speculative bulk prepayment. DeepSeek explicitly recommends topping up according to actual usage.
For a managed open-model control group, compare Novita’s current DeepSeek and open-model routes. Confirm the exact checkpoint and rate before assuming it reproduces the first-party endpoint.
Affiliate disclosure: AI Pricing Guru may earn a commission from the sponsored Novita link at no extra cost to you. It does not affect this analysis.
Labs status: both DeepSeek models are included
DeepSeek V4 Flash and V4 Pro are already represented in AI Pricing Guru Labs. In the latest accepted 49-task run, each scored 48/49 with zero API errors. The Flash row remains explicitly labeled pre-0731 because the launch-day refresh failed for insufficient credit; we rejected that incomplete run instead of mixing it into the leaderboard.
This warning changes no model artifact, endpoint, or active rate, so a fresh inference run would not answer a new capability question. Once DeepSeek publishes and activates the new prices, we will recalculate both cost-per-correct figures against the accepted token counts and rerun only if the served model or endpoint also changes.
Bottom line
DeepSeek’s warning is now resolved: the provider has published peak and off-peak V4 rates that start on August 16. Use our current pricing-update analysis for the generated rate table and buyer recommendations.
For model-specific context, read our DeepSeek V4 Flash 0731 analysis and the broader DeepSeek API pricing guide.
Sources: DeepSeek’s official models and pricing documentation, checked August 6, 2026, the Hacker News discussion that surfaced the warning, and DeepSeek’s later V4 Pro and pricing announcement.