Ask
Price change

DeepSeek confirms peak/off-peak pricing, effective Aug 16, 2026

DeepSeek pricing

DeepSeek confirmed peak/off-peak API pricing effective 2026-08-16: peak rates run roughly 3-12x today's flat prices, off-peak rates roughly 1.5-2.5x.

Before

Flat per-token rates with an unspecified 'significant increase expected' footnote (first seen 2026-08-11)

After

Peak/off-peak billing: V4-Flash cache-miss input $0.22 off-peak / $0.44 peak (was $0.14 flat); V4-Pro cache-miss input $0.66 off-peak / $1.32 peak (was $0.435 flat); peak hours 01:00-04:00 and 06:00-10:00 UTC

Three days after DeepSeek’s Models & Pricing page first warned of an unspecified “significant” price increase, the company published the actual rate card. Starting at 16:00 UTC on August 16, 2026, DeepSeek API pricing splits into peak and off-peak windows — peak hours are 01:00–04:00 and 06:00–10:00 UTC, with all other hours off-peak at exactly half the peak rate.

The increases are steep and uneven across line items. DeepSeek-V4-Flash cache-miss input rises from a flat $0.14/1M to $0.22/1M off-peak and $0.44/1M peak; output rises from $0.28/1M to $0.66/1M off-peak and $1.32/1M peak. The biggest relative jump is on cache-hit input — DeepSeek’s cheapest and most aggressively marketed rate — which moves from $0.0028/1M to $0.007/1M off-peak (2.5x) and $0.014/1M peak (5x) on V4-Flash, and from $0.003625/1M to $0.022/1M off-peak and $0.044/1M peak (over 12x) on the higher-capacity V4-Pro model. No grace period or legacy-rate opt-out has been published.

The same capture also showed DeepSeek-V4-Pro reaching feature parity with V4-Flash on the Responses API (previously V4-Pro-only support was pending) and picking up a dated build tag, DeepSeek-V4-Pro-0813.

From DeepSeek's pricing timeline
DeepSeek Announces Peak/Off-Peak Pricing — Effective Aug 16, 2026

DeepSeek replaced its vague "significant price increase" footnote with a concrete time-of-use rate card: peak-hour rates (01:00-04:00 and 06:00-10:00 UTC) roughly 3-11x current per-token prices depending on the line item, with off-peak rates at half of peak, effective 16:00 UTC on 2026-08-16. V4-Flash cache-miss input moves from a flat $0.14 to $0.22 off-peak / $0.44 peak; V4-Pro cache-miss input moves from a flat $0.435 to $0.66 off-peak / $1.32 peak. DeepSeek-V4-Pro also picked up a dated build suffix (-0813) and gained Responses API support, reaching feature parity with V4-Flash.

About DeepSeek
api-docs.deepseek.com ↗

DeepSeek offers a free web chat product and a pay-per-token API, billed since 2026-08-16 on a confirmed peak/off-peak rate card: DeepSeek-V4-Flash (general-purpose, from $0.007/1M cache-hit input off-peak / $0.014 peak, $0.22 cache-miss in off-peak / $0.44 peak, $0.66 out off-peak / $1.32 peak), DeepSeek-V4-Pro (higher-capacity, from $0.022/1M cache-hit input off-peak / $0.044 peak), and the experimental DeepSeek-V4-Flash-Vision-Exp (billed on V4-Flash's rate card) — all with a 1M-token context window and dramatically cheaper than equivalent OpenAI and Anthropic models.

Pricing model freemiumpure usage
Billing units tokensapi calls
Sales motion self serveplg
Free tier
Yes
Commits
None
Transparency
public

DeepSeek pricing history

  1. Aug 2026
    DeepSeek Announces Peak/Off-Peak Pricing — Effective Aug 16, 2026
  2. Mar 2025
    DeepSeek-V3-0324 Update — Improved Coding
  3. Jan 2025
    Nvidia Stock Drops 17% Following R1 Release
  4. Jan 2025
    DeepSeek-R1 Released — Reasoning Model, MIT Open-Source
  5. Dec 2024
    DeepSeek-V3 Released — Frontier Performance at $0.27/1M
Full DeepSeek timeline

More DeepSeek activity

All pricing activity