Gemini 3.1 Flash-Lite
gemini-3.1-flash-liteDeveloper API · Standard · text/image/videoGemini Developer API
- Input
- $0.25
- Cache read
- $0.025
- Output
- $1.50
- Context
- —
Compare Gemini API model tiers, cached-input rates, output rates, introductory dates, and explicit cache-storage caveats.
Enter monthly workload, tokens, and quality pass rates. The calculator resolves eligible price bands, then shows cost per accepted answer and the candidate break-even pass rate.
Candidate break-even: 59.6% pass rate. Your scenario assumes 85%.
Warm cache-read scenario, not general ROI. It excludes cache writes/storage, tools, regional uplifts, retries, and quality failures. A dash means no published cache-read rate.
Gemini API pricing ↗Gemini Developer API and Vertex AI are not interchangeable billing surfaces. These cards retain the documented surface, tier, and effective dates instead of presenting one blended Gemini price.
Rates are USD per one million tokens. Future and transitional cards remain labeled with their effective dates instead of silently replacing today’s tier.
| Provider / model | Price scope | Input USD / 1M | Cache read USD / 1M | Output USD / 1M | Context | Source |
|---|---|---|---|---|---|---|
GoogleGemini 3.1 Flash-Litegemini-3.1-flash-lite | Developer API · Standard · text/image/videoGemini Developer API | $0.25 | $0.025 | $1.50 | — | Official ↗ |
GoogleGemini 3.1 Pro Previewgemini-3.1-pro-preview | Developer API · Standard · over 200K promptGemini Developer API | $4 | $0.4 | $18 | — | Official ↗ |
GoogleGemini 3.1 Pro Previewgemini-3.1-pro-preview | Developer API · Standard · up to 200K promptGemini Developer API | $2 | $0.2 | $12 | — | Official ↗ |
GoogleGemini 3.5 Flashgemini-3.5-flash | Developer API · StandardGemini Developer API | $1.50 | $0.15 | $9 | — | Official ↗ |
GoogleGemini 3.5 Flash-Litegemini-3.5-flash-lite | Developer API · StandardGemini Developer API | $0.3 | $0.03 | $2.50 | — | Official ↗ |
GoogleGemini 3.6 Flashgemini-3.6-flash | Developer API · Standard · introductoryGemini Developer APIThrough 31 Dec 2026 | $0.75 | $0.075 | $3.75 | — | Official ↗ |
GoogleGemini 3.7 Flashgemini-3.7-flash | Developer API · Standard · introductoryGemini Developer APIThrough 31 Dec 2026 | $0.75 | $0.075 | $3.75 | — | Official ↗ |
Rates in USD per 1M tokens
gemini-3.1-flash-liteDeveloper API · Standard · text/image/videoGemini Developer API
gemini-3.1-pro-previewDeveloper API · Standard · over 200K promptGemini Developer API
gemini-3.1-pro-previewDeveloper API · Standard · up to 200K promptGemini Developer API
gemini-3.5-flashDeveloper API · StandardGemini Developer API
gemini-3.5-flash-liteDeveloper API · StandardGemini Developer API
gemini-3.6-flashDeveloper API · Standard · introductoryGemini Developer API
gemini-3.7-flashDeveloper API · Standard · introductoryGemini Developer API
USD per 1M tokens. Snapshot verified 2026-08-16. Cache writes, cache storage, tools, regions, and provider-specific thresholds may be billed separately.
A missing cache rate is shown as “—”, never treated as free. Consumer chat-plan quotas are not API prices.
Gemini Developer API and Vertex AI are not interchangeable billing surfaces. These cards retain the documented surface, tier, and effective dates instead of presenting one blended Gemini price.
Count retries, fallbacks, tool calls, latency failures, and answers rejected by the quality rubric. A lower estimated token bill is useful only when the workload still succeeds.
Open the controlled A/B labThis page uses TokenGauge’s 2026-08-16 official-source snapshot. Follow the linked provider pages before making a production commitment because prices can change.
No. It models input, output, and a warm cache-read share. Cache writes or storage, tools, retries, regional uplifts, taxes, batch modes, and quality failures can change the final invoice.
No. ChatGPT, Claude, Gemini, Grok, Kimi, and other consumer-plan quotas are separate from provider API token billing.