DeepSeek V4 Flash
deepseek-v4-flashOff-peak · scheduled UTC windowsGlobal
- Input
- $0.22
- Cache read
- $0.007
- Output
- $0.66
- Context
- 1M
Compare DeepSeek API token rates with effective dates and time-sensitive pricing rules kept visible.
Enter monthly workload, tokens, and quality pass rates. The calculator resolves eligible price bands, then shows cost per accepted answer and the candidate break-even pass rate.
Candidate break-even: 55.9% pass rate. Your scenario assumes 85%.
Warm cache-read scenario, not general ROI. It excludes cache writes/storage, tools, regional uplifts, retries, and quality failures. A dash means no published cache-read rate.
DeepSeek API pricing ↗DeepSeek can publish scheduled and time-window pricing changes. TokenGauge keeps effective boundaries on separate rate cards and never applies a future rate before it begins.
Rates are USD per one million tokens. Future and transitional cards remain labeled with their effective dates instead of silently replacing today’s tier.
| Provider / model | Price scope | Input USD / 1M | Cache read USD / 1M | Output USD / 1M | Context | Source |
|---|---|---|---|---|---|---|
DeepSeekDeepSeek V4 Flashdeepseek-v4-flash | Off-peak · scheduled UTC windowsGlobalFrom 16 Aug 2026 | $0.22 | $0.007 | $0.66 | 1M | Official ↗ |
DeepSeekDeepSeek V4 Flashdeepseek-v4-flash | Peak · 01:00–04:00 and 06:00–10:00 UTCGlobalFrom 16 Aug 2026 | $0.44 | $0.014 | $1.32 | 1M | Official ↗ |
DeepSeekDeepSeek V4 Flashdeepseek-v4-flash | Standard · through 2026-08-16 16:00 UTCGlobalThrough 16 Aug 2026 | $0.14 | $0.0028 | $0.28 | 1M | Official ↗ |
DeepSeekDeepSeek V4 Prodeepseek-v4-pro | Off-peak · scheduled UTC windowsGlobalFrom 16 Aug 2026 | $0.66 | $0.022 | $1.98 | 1M | Official ↗ |
DeepSeekDeepSeek V4 Prodeepseek-v4-pro | Peak · 01:00–04:00 and 06:00–10:00 UTCGlobalFrom 16 Aug 2026 | $1.32 | $0.044 | $3.96 | 1M | Official ↗ |
DeepSeekDeepSeek V4 Prodeepseek-v4-pro | Standard · through 2026-08-16 16:00 UTCGlobalThrough 16 Aug 2026 | $0.435 | $0.003625 | $0.87 | 1M | Official ↗ |
Rates in USD per 1M tokens
deepseek-v4-flashOff-peak · scheduled UTC windowsGlobal
deepseek-v4-flashPeak · 01:00–04:00 and 06:00–10:00 UTCGlobal
deepseek-v4-flashStandard · through 2026-08-16 16:00 UTCGlobal
deepseek-v4-proOff-peak · scheduled UTC windowsGlobal
deepseek-v4-proPeak · 01:00–04:00 and 06:00–10:00 UTCGlobal
deepseek-v4-proStandard · through 2026-08-16 16:00 UTCGlobal
USD per 1M tokens. Snapshot verified 2026-08-16. Cache writes, cache storage, tools, regions, and provider-specific thresholds may be billed separately.
A missing cache rate is shown as “—”, never treated as free. Consumer chat-plan quotas are not API prices.
DeepSeek can publish scheduled and time-window pricing changes. TokenGauge keeps effective boundaries on separate rate cards and never applies a future rate before it begins.
Count retries, fallbacks, tool calls, latency failures, and answers rejected by the quality rubric. A lower estimated token bill is useful only when the workload still succeeds.
Open the controlled A/B labThis page uses TokenGauge’s 2026-08-16 official-source snapshot. Follow the linked provider pages before making a production commitment because prices can change.
No. It models input, output, and a warm cache-read share. Cache writes or storage, tools, retries, regional uplifts, taxes, batch modes, and quality failures can change the final invoice.
No. ChatGPT, Claude, Gemini, Grok, Kimi, and other consumer-plan quotas are separate from provider API token billing.