LLM API

Google Gemini API pricing

Useful for Google-centric teams evaluating Gemini pricing across context tiers, modalities, caching, and batch terms.

Decision summary: Choose Gemini when Google ecosystem fit matters; account for thinking-token output pricing, modality rates, cache storage, and >200k prompt tiers.
Pricing checked 2026-08-26high confidence

Pricing overview

Pricing varies by model, modality, context tier, token type, context caching, and batch usage.

Source-tracked model data

Google Gemini API model prices in one table

Example cost uses 1,000,000 input tokens and 300,000 output tokens at each model's default tracked rate, with no cache discount.

ModelInput priceCached inputOutput priceContextExample cost
Gemini 3.6 Flash
Stable
$1.50 / 1M$0.15 / 1M$7.50 / 1M1M tokens$3.75
Gemini 3.5 Flash-Lite
Stable
$0.30 / 1M$0.03 / 1M$2.50 / 1M1M tokens$1.05
Gemini 3.5 Flash
Stable
$1.50 / 1M$0.15 / 1M$9.00 / 1M1M tokens$4.20
Gemini 3.5 Live Translate Preview
Preview
$0.25 / 1M$0.025 / 1M$1.50 / 1M131.1K tokens$0.70
Gemini 3.1 Flash-Lite
Stable
$0.25 / 1M$0.025 / 1M$1.50 / 1M1M tokens$0.70
Gemini 3.1 Pro Preview
Preview
$2.00 / 1M$0.20 / 1M$12.00 / 1M1M tokens$5.60
Gemini 3 Flash Preview
Preview
$0.50 / 1M$0.05 / 1M$3.00 / 1M1M tokens$1.40
Gemini 2.5 Flash
Stable
$0.30 / 1M$0.03 / 1M$2.50 / 1M1M tokens$1.05
Gemini 2.5 Flash-Lite
Stable
$0.10 / 1M$0.01 / 1M$0.40 / 1M1M tokens$0.22
Gemini 2.5 Pro
Stable
$1.25 / 1M$0.125 / 1M$10.00 / 1M1M tokens$4.25

Cost boundary: this example compares token charges only. It does not assume equal output quality, latency, retry rates, or output length across models.

Official price history

Google Gemini API recorded model price changes

Verified changes for exact models and pricing profiles available through this provider. New model generations are not treated as price changes.

Google · Standard · Text

Gemini 2.5 Flash

Effective 2025-06-17
RateEarlier priceLater recorded priceChangePercentage
Input / 1M$0.15$0.30+$0.15+100.0%
Output / 1M$3.50$2.50-$1.00-28.6%
Checked 2026-07-12Official source ↗

API snapshot: gemini-2.5-flash-preview-04-17gemini-2.5-flash

Google released the stable Gemini 2.5 Flash API snapshot with one output price that includes thinking tokens. This is a preview-to-stable version transition, not a like-for-like price change across every request mode.

Google · Standard · Short Context

Gemini 1.5 Pro

Effective 2024-10-01
RateEarlier priceLater recorded priceChangePercentage
Input / 1M$3.50$1.25-$2.25-64.3%
Output / 1M$10.50$5.00-$5.50-52.4%
Checked 2026-07-17Official source ↗

API snapshot: gemini-1.5-pro-001gemini-1.5-pro-002

Google reduced Gemini 1.5 Pro prices for short- and long-context prompts.

Google · Standard · Long Context

Gemini 1.5 Pro

Effective 2024-10-01
RateEarlier priceLater recorded priceChangePercentage
Input / 1M$7.00$2.50-$4.50-64.3%
Output / 1M$21.00$10.00-$11.00-52.4%
Checked 2026-07-17Official source ↗

API snapshot: gemini-1.5-pro-001gemini-1.5-pro-002

Google reduced Gemini 1.5 Pro prices for short- and long-context prompts.

Google · Standard · Short Context

Gemini 1.5 Flash

Effective 2024-08-12
RateEarlier priceLater recorded priceChangePercentage
Input / 1M$0.35$0.075-$0.275-78.6%
Cached input / 1M$0.0875$0.0188-$0.06875-78.6%
Output / 1M$1.05$0.30-$0.75-71.4%
Checked 2026-07-17Official source ↗

API snapshot: gemini-1.5-flash-001gemini-1.5-flash-002

Google reduced Gemini 1.5 Flash prices across short-context, long-context, and cached-input tiers.

Google · Standard · Long Context

Gemini 1.5 Flash

Effective 2024-08-12
RateEarlier priceLater recorded priceChangePercentage
Input / 1M$0.70$0.15-$0.55-78.6%
Cached input / 1M$0.175$0.0375-$0.1375-78.6%
Output / 1M$2.10$0.60-$1.50-71.4%
Checked 2026-07-17Official source ↗

API snapshot: gemini-1.5-flash-001gemini-1.5-flash-002

Google reduced Gemini 1.5 Flash prices across short-context, long-context, and cached-input tiers.

What affects cost

  • Prompt size above or below 200k tokens
  • Output tokens including thinking tokens
  • Audio input rates
  • Context caching and cache storage
  • Batch terms

Best for

  • Google ecosystem teams
  • multimodal products
  • workloads that can benefit from lower-cost Flash-Lite routing

Not ideal for

  • teams that cannot track modality-specific terms
  • buyers needing one flat rate across all prompt sizes
  • workloads where thinking-token output needs strict caps
FAQ

Google Gemini API pricing questions

Does Gemini 2.5 Pro change price above 200k prompt tokens?

Yes. StackLens stores separate <=200k and >200k standard paid tiers, and the calculator switches tiers above 200k input tokens per request.

Does Gemini output pricing include thinking tokens?

Yes for the tracked Gemini values in this dataset. Output estimates should account for generated answer length and thinking tokens.