Codestral
codestral-latestStandardGlobal
- Input
- $0.3
- Cache read
- $0.03
- Output
- $0.9
- Context
- —
Estimate Mistral API input and output token costs from a dated, official-source pricing snapshot.
Enter monthly workload, tokens, and quality pass rates. The calculator resolves eligible price bands, then shows cost per accepted answer and the candidate break-even pass rate.
Candidate break-even: 59.6% pass rate. Your scenario assumes 85%.
Warm cache-read scenario, not general ROI. It excludes cache writes/storage, tools, regional uplifts, retries, and quality failures. A dash means no published cache-read rate.
Mistral API pricing ↗A missing cache rate is displayed as unavailable, not zero. The calculator falls back to the normal input price whenever no separate cache-read rate is published.
Rates are USD per one million tokens. Future and transitional cards remain labeled with their effective dates instead of silently replacing today’s tier.
| Provider / model | Price scope | Input USD / 1M | Cache read USD / 1M | Output USD / 1M | Context | Source |
|---|---|---|---|---|---|---|
Mistral AICodestralcodestral-latest | StandardGlobal | $0.3 | $0.03 | $0.9 | — | Official ↗ |
Mistral AIMinistral 3 14Bministral-14b-latest | StandardGlobal | $0.2 | $0.02 | $0.2 | — | Official ↗ |
Mistral AIMinistral 3 3Bministral-3b-latest | StandardGlobal | $0.1 | $0.01 | $0.1 | — | Official ↗ |
Mistral AIMinistral 3 8Bministral-8b-latest | StandardGlobal | $0.15 | $0.015 | $0.15 | — | Official ↗ |
Mistral AIMistral Large 3mistral-large-latest | StandardGlobal | $0.5 | $0.05 | $1.50 | — | Official ↗ |
Mistral AIMistral Medium 3.5mistral-medium-latest | StandardGlobal | $1.50 | $0.15 | $7.50 | — | Official ↗ |
Mistral AIMistral Small 4mistral-small-latest | StandardGlobal | $0.15 | $0.015 | $0.6 | — | Official ↗ |
Rates in USD per 1M tokens
codestral-latestStandardGlobal
ministral-14b-latestStandardGlobal
ministral-3b-latestStandardGlobal
ministral-8b-latestStandardGlobal
mistral-large-latestStandardGlobal
mistral-medium-latestStandardGlobal
mistral-small-latestStandardGlobal
USD per 1M tokens. Snapshot verified 2026-08-16. Cache writes, cache storage, tools, regions, and provider-specific thresholds may be billed separately.
A missing cache rate is shown as “—”, never treated as free. Consumer chat-plan quotas are not API prices.
A missing cache rate is displayed as unavailable, not zero. The calculator falls back to the normal input price whenever no separate cache-read rate is published.
Count retries, fallbacks, tool calls, latency failures, and answers rejected by the quality rubric. A lower estimated token bill is useful only when the workload still succeeds.
Open the controlled A/B labThis page uses TokenGauge’s 2026-08-16 official-source snapshot. Follow the linked provider pages before making a production commitment because prices can change.
No. It models input, output, and a warm cache-read share. Cache writes or storage, tools, retries, regional uplifts, taxes, batch modes, and quality failures can change the final invoice.
No. ChatGPT, Claude, Gemini, Grok, Kimi, and other consumer-plan quotas are separate from provider API token billing.