GOOGLE API PRICING

Google Gemini token cost calculator.

Compare Gemini API model tiers, cached-input rates, output rates, introductory dates, and explicit cache-storage caveats.

Models
6
Rate cards
7
Input range
$0.25$4
Output range
$1.50$18
Cache-priced cards
7
Verified
2026-08-16
GOOGLE WORKLOAD

Model the calls you actually make.

Enter monthly workload, tokens, and quality pass rates. The calculator resolves eligible price bands, then shows cost per accepted answer and the candidate break-even pass rate.

Estimated monthly API spend · Google
Gemini 3.7 FlashBefore · Developer API · Standard · introductory: $0.75 input · $0.075 cache read · $3.75 output / 1MAfter · Developer API · Standard · introductory: $0.75 input · $0.075 cache read · $3.75 output / 1MGemini Developer API · price bands selected automatically from input size
Before$41.25
After$27.32
Potential saving: $13.93 (33.8%)
QUALITY-ADJUSTED COST
Before / accepted answer$0.004690% pass assumption
After / accepted answer$0.003285% pass assumption

Candidate break-even: 59.6% pass rate. Your scenario assumes 85%.

Warm cache-read scenario, not general ROI. It excludes cache writes/storage, tools, regional uplifts, retries, and quality failures. A dash means no published cache-read rate.

Gemini API pricing
NEXT STEPValidate the 29.9% accepted-answer advantage.

The estimate is not a saving until the same workload still passes its quality bar. Compare a supported recipe in the lab, or use the evidence library to design a provider-specific test.

Gemini Developer API and Vertex AI are not interchangeable billing surfaces. These cards retain the documented surface, tier, and effective dates instead of presenting one blended Gemini price.

DATED RATE CARDS

Google prices with their scope intact.

Rates are USD per one million tokens. Future and transitional cards remain labeled with their effective dates instead of silently replacing today’s tier.

Showing 7 of 52 rate cards

Official provider API rate cards, USD per one million tokens
Provider / modelPrice scopeInput
USD / 1M
Cache read
USD / 1M
Output
USD / 1M
ContextSource
GoogleGemini 3.1 Flash-Litegemini-3.1-flash-liteDeveloper API · Standard · text/image/videoGemini Developer API$0.25$0.025$1.50Official ↗
GoogleGemini 3.1 Pro Previewgemini-3.1-pro-previewDeveloper API · Standard · over 200K promptGemini Developer API$4$0.4$18Official ↗
GoogleGemini 3.1 Pro Previewgemini-3.1-pro-previewDeveloper API · Standard · up to 200K promptGemini Developer API$2$0.2$12Official ↗
GoogleGemini 3.5 Flashgemini-3.5-flashDeveloper API · StandardGemini Developer API$1.50$0.15$9Official ↗
GoogleGemini 3.5 Flash-Litegemini-3.5-flash-liteDeveloper API · StandardGemini Developer API$0.3$0.03$2.50Official ↗
GoogleGemini 3.6 Flashgemini-3.6-flashDeveloper API · Standard · introductoryGemini Developer APIThrough 31 Dec 2026$0.75$0.075$3.75Official ↗
GoogleGemini 3.7 Flashgemini-3.7-flashDeveloper API · Standard · introductoryGemini Developer APIThrough 31 Dec 2026$0.75$0.075$3.75Official ↗

Rates in USD per 1M tokens

Google

Gemini 3.1 Flash-Lite

gemini-3.1-flash-lite

Developer API · Standard · text/image/videoGemini Developer API

Input
$0.25
Cache read
$0.025
Output
$1.50
Context
Open official source in a new tab
Google

Gemini 3.1 Pro Preview

gemini-3.1-pro-preview

Developer API · Standard · over 200K promptGemini Developer API

Input
$4
Cache read
$0.4
Output
$18
Context
Open official source in a new tab
Google

Gemini 3.1 Pro Preview

gemini-3.1-pro-preview

Developer API · Standard · up to 200K promptGemini Developer API

Input
$2
Cache read
$0.2
Output
$12
Context
Open official source in a new tab
Google

Gemini 3.6 Flash

gemini-3.6-flash

Developer API · Standard · introductoryGemini Developer API

Input
$0.75
Cache read
$0.075
Output
$3.75
Context
Effective through 31 Dec 2026Open official source in a new tab
Google

Gemini 3.7 Flash

gemini-3.7-flash

Developer API · Standard · introductoryGemini Developer API

Input
$0.75
Cache read
$0.075
Output
$3.75
Context
Effective through 31 Dec 2026Open official source in a new tab

USD per 1M tokens. Snapshot verified 2026-08-16. Cache writes, cache storage, tools, regions, and provider-specific thresholds may be billed separately.

A missing cache rate is shown as “—”, never treated as free. Consumer chat-plan quotas are not API prices.

MEASUREMENT RULE

Price accepted answers, not token deltas

Count retries, fallbacks, tool calls, latency failures, and answers rejected by the quality rubric. A lower estimated token bill is useful only when the workload still succeeds.

Open the controlled A/B lab
QUESTIONS

Before you trust the estimate.

How current are the Google API prices?

This page uses TokenGauge’s 2026-08-16 official-source snapshot. Follow the linked provider pages before making a production commitment because prices can change.

Does this calculator include every Google charge?

No. It models input, output, and a warm cache-read share. Cache writes or storage, tools, retries, regional uplifts, taxes, batch modes, and quality failures can change the final invoice.

Are consumer chat subscriptions included?

No. ChatGPT, Claude, Gemini, Grok, Kimi, and other consumer-plan quotas are separate from provider API token billing.