Grok 4.20 Reasoning
grok-4.20-0309-reasoningStandard · below 200K promptxAI direct API
- Input
- $1.25
- Cache read
- $0.2
- Output
- $2.50
- Context
- 1M
Estimate Grok API spend from current official input, cached-input, output, context-window, and regional price scopes.
Enter monthly workload, tokens, and quality pass rates. The calculator resolves eligible price bands, then shows cost per accepted answer and the candidate break-even pass rate.
Candidate break-even: 58.9% pass rate. Your scenario assumes 85%.
Warm cache-read scenario, not general ROI. It excludes cache writes/storage, tools, regional uplifts, retries, and quality failures. A dash means no published cache-read rate.
xAI model pricing ↗A published cache-read discount is not a guaranteed cache hit. Model realistic warm-cache shares and measure actual provider usage before treating a scenario as savings.
Rates are USD per one million tokens. Future and transitional cards remain labeled with their effective dates instead of silently replacing today’s tier.
| Provider / model | Price scope | Input USD / 1M | Cache read USD / 1M | Output USD / 1M | Context | Source |
|---|---|---|---|---|---|---|
xAIGrok 4.20 Reasoninggrok-4.20-0309-reasoning | Standard · below 200K promptxAI direct API | $1.25 | $0.2 | $2.50 | 1M | Official ↗ |
xAIGrok 4.3grok-4.3 | Standard · 200K+ promptxAI direct API | $2.50 | $0.4 | $5 | 1M | Official ↗ |
xAIGrok 4.3grok-4.3 | Standard · below 200K promptxAI direct API | $1.25 | $0.2 | $2.50 | 1M | Official ↗ |
xAIGrok 4.5grok-4.5 | Standard · below 200K promptxAI direct API | $2 | $0.3 | $6 | 500K | Official ↗ |
xAIGrok 4.6grok-4.6 | Standard · 200K+ promptxAI direct API | $4 | $1 | $12 | 500K | Official ↗ |
xAIGrok 4.6grok-4.6 | Standard · below 200K promptxAI direct API | $2 | $0.5 | $6 | 500K | Official ↗ |
xAIGrok Build 0.1grok-build-0.1 | Standard · below 200K promptxAI direct API | $1 | $0.2 | $2 | 256K | Official ↗ |
Rates in USD per 1M tokens
grok-4.20-0309-reasoningStandard · below 200K promptxAI direct API
grok-4.3Standard · 200K+ promptxAI direct API
grok-4.3Standard · below 200K promptxAI direct API
grok-4.5Standard · below 200K promptxAI direct API
grok-4.6Standard · 200K+ promptxAI direct API
grok-4.6Standard · below 200K promptxAI direct API
grok-build-0.1Standard · below 200K promptxAI direct API
USD per 1M tokens. Snapshot verified 2026-08-16. Cache writes, cache storage, tools, regions, and provider-specific thresholds may be billed separately.
A missing cache rate is shown as “—”, never treated as free. Consumer chat-plan quotas are not API prices.
A published cache-read discount is not a guaranteed cache hit. Model realistic warm-cache shares and measure actual provider usage before treating a scenario as savings.
Count retries, fallbacks, tool calls, latency failures, and answers rejected by the quality rubric. A lower estimated token bill is useful only when the workload still succeeds.
Open the controlled A/B labThis page uses TokenGauge’s 2026-08-16 official-source snapshot. Follow the linked provider pages before making a production commitment because prices can change.
No. It models input, output, and a warm cache-read share. Cache writes or storage, tools, retries, regional uplifts, taxes, batch modes, and quality failures can change the final invoice.
No. ChatGPT, Claude, Gemini, Grok, Kimi, and other consumer-plan quotas are separate from provider API token billing.