DeepSeek V4 Flash
Off-peak · scheduled UTC windows · Global- Monthly API spend
- $9.90
- Cost / accepted answer
- $0.0011
- Quality assumption
- 90%
$0.22 input · $0.007 cache read · $0.66 output / 1M
Official source ↗Compare DeepSeek and Moonshot Kimi API spend without hiding effective dates, cache modes, context tiers, or batch-only discounts.
Both sides receive the same requests and token counts. Set separate quality pass rates so a low token price cannot hide rejected answers.
$0.22 input · $0.007 cache read · $0.66 output / 1M
Official source ↗$3 input · $0.3 cache read · $15 output / 1M
Official source ↗At the entered pass rates, its estimated cost per accepted answer is $0.0011. This is scenario math, not evidence that the selected models have equal capabilities.
API rates are USD per one million tokens. The estimate excludes cache writes or storage, tools, retries, regional uplifts, taxes, latency failures, and consumer-plan quotas.
DeepSeek time-window prices and Kimi batch or cache tiers apply only under their documented conditions. A scheduled or batch rate is not a general on-demand price.
4 currently effective cards are available in the calculator. Context bands and regions remain separate.
Open DeepSeek pricing4 currently effective cards are available in the calculator. A missing cache rate is never treated as free.
Open Kimi / Moonshot AI pricingThere is no provider-wide answer. Cost depends on the exact model tier, input and output size, cache behavior, region, execution mode, and the share of answers that pass the workload’s quality bar.
No. The calculator keeps separate quality-pass assumptions so you can enter measured results. Equal token counts or prices do not establish equivalent capabilities.
No. ChatGPT, Claude, Gemini, Grok, Kimi, and other consumer-plan quotas are separate from API token billing.