Polarison

Gemini API pricing

Google charges per million tokens, with a lower rate for the tokens you send than for the tokens the model writes. The 6 Gemini models we track run from $0.30 / $2.50 (Gemini 3.5 Flash-Lite) to $2.00 / $12.00 (Gemini 3.1 Pro).

Prices from Google’s official pricing page, verified October 3, 2026.

Every Gemini model, cheapest first

USD per 1M tokens, standard tier.

ModelInputCached inputOutputLong contextBatchContext
Gemini 3.5 Flash-Lite
gemini-3.5-flash-lite
$0.30$0.03$2.50—50% off1.05M tokens
Gemini 2.5 Flash
gemini-2.5-flash
$0.30$0.03$2.50——1.05M tokens
Gemini 3.8 Flash
gemini-3.8-flash
$0.75$0.075$3.75—50% off1.05M tokens
Gemini 3.5 Flash
gemini-3.5-flash
$1.50$0.15$9.00—50% off1.05M tokens
Gemini 2.5 Pro
gemini-2.5-pro
$1.25$0.125$10.00$2.50 in / $15.00 out above 200K prompt tokens—1.05M tokens
Gemini 3.1 ProPreview
gemini-3.1-pro-preview
$2.00$0.20$12.00$4.00 in / $18.00 out above 200K prompt tokens—1.05M tokens

Gemini API cost calculator

Standard-tier list prices. Long-context rates apply automatically where the provider publishes a threshold. Excludes cache-write surcharges, taxes, batch and volume discounts.

Pricing details by model

Gemini 3.5 Flash-Lite
  • The same input price applies to text, image, video and audio.
  • Output price includes thinking tokens.
Gemini 2.5 Flash
  • Audio input costs $1.00 per 1M tokens; text, image and video cost $0.30.
  • Output price includes thinking tokens.
Gemini 3.8 Flash
  • Introductory price through December 31, 2026. From January 1, 2027: $1.50 input, $0.15 cached, $7.50 output per 1M tokens.
  • Output price includes thinking tokens.
  • Priority inference costs 1.8x the standard rate.
Gemini 3.5 Flash
  • Output price includes thinking tokens.
  • Context cache storage costs $1.00 per 1M tokens per hour.
  • Priority inference costs 1.8x the standard rate.
Gemini 2.5 Pro
  • Prompts over 200K tokens are billed at $2.50 input / $15 output per 1M tokens.
  • Output price includes thinking tokens.
Gemini 3.1 Pro
  • Prompts over 200K tokens are billed at $4 input / $18 output per 1M tokens.
  • No free tier on the Gemini API.
  • Output price includes thinking tokens.

Recent Gemini pricing changes

  • Gemini 3.8 Flash introductory pricing ends

    The standard price rises from $0.75 to $1.50 per 1M input tokens, from $0.075 to $0.15 per 1M cached input tokens and from $3.75 to $7.50 per 1M output tokens.

  • Google’s Gemini 3.8 text-to-speech prices double

    Gemini 3.8 Flash TTS goes from $0.50 to $1.00 per 1M text input tokens and from $9.00 to $18.00 per 1M audio output tokens (about $0.0135 to $0.027 per minute of audio). Gemini 3.8 Flash-Lite TTS goes from $6.00 to $12.00 per 1M audio output tokens. Batch, Flex and Priority rates double too.

All pricing changes

Subscription instead of the API?

These plans run Gemini models for a flat monthly price. See whether one costs less than paying per token with the subscription vs API calculator.

Frequently asked questions

How much does the Gemini API cost?
Gemini API prices range from $0.30 input / $2.50 output per 1M tokens for Gemini 3.5 Flash-Lite to $2.00 / $12.00 for Gemini 3.1 Pro. A support chatbot handling 100,000 replies a month costs about $145.00 on Gemini 3.5 Flash-Lite.
What is the cheapest Gemini model?
Gemini 3.5 Flash-Lite, at a blended $0.85 per 1M tokens (3 input : 1 output) — 81% less per token than Gemini 3.1 Pro.
Does the Gemini API have a batch discount?
Yes. Batch requests, which return results within a window of up to 24 hours instead of immediately, cost 50% less on the Gemini models that support it.
Which Gemini model is the most capable?
Gemini 3.8 Flash has the highest Epoch Capabilities Index of the Gemini models we track (156.5), at $0.75 input / $3.75 output per 1M tokens.