Gemini API pricing
Google charges per million tokens, with a lower rate for the tokens you send than for the tokens the model writes. The 6 Gemini models we track run from $0.30 / $2.50 (Gemini 3.5 Flash-Lite) to $2.00 / $12.00 (Gemini 3.1 Pro).
Prices from Google’s official pricing page, verified October 3, 2026.
Every Gemini model, cheapest first
USD per 1M tokens, standard tier.
| Model | Input | Cached input | Output | Long context | Batch | Context |
|---|---|---|---|---|---|---|
| Gemini 3.5 Flash-Lite gemini-3.5-flash-lite | $0.30 | $0.03 | $2.50 | — | 50% off | 1.05M tokens |
| Gemini 2.5 Flash gemini-2.5-flash | $0.30 | $0.03 | $2.50 | — | — | 1.05M tokens |
| Gemini 3.8 Flash gemini-3.8-flash | $0.75 | $0.075 | $3.75 | — | 50% off | 1.05M tokens |
| Gemini 3.5 Flash gemini-3.5-flash | $1.50 | $0.15 | $9.00 | — | 50% off | 1.05M tokens |
| Gemini 2.5 Pro gemini-2.5-pro | $1.25 | $0.125 | $10.00 | $2.50 in / $15.00 out above 200K prompt tokens | — | 1.05M tokens |
| Gemini 3.1 ProPreview gemini-3.1-pro-preview | $2.00 | $0.20 | $12.00 | $4.00 in / $18.00 out above 200K prompt tokens | — | 1.05M tokens |
Gemini API cost calculator
- $195.00/mo$0.0039 / request
- Gemini 2.5 FlashGoogle$195.00/mo$0.0039 / request
- Gemini 3.8 FlashGoogle$412.50/mo$0.0083 / request
- Gemini 2.5 ProGoogle$800.00/mo$0.02 / request
- Gemini 3.5 FlashGoogle$870.00/mo$0.02 / request
- Gemini 3.1 ProGoogle$1,160/mo$0.02 / request
Standard-tier list prices. Long-context rates apply automatically where the provider publishes a threshold. Excludes cache-write surcharges, taxes, batch and volume discounts.
Pricing details by model
- Gemini 3.5 Flash-Lite
- The same input price applies to text, image, video and audio.
- Output price includes thinking tokens.
- Gemini 2.5 Flash
- Audio input costs $1.00 per 1M tokens; text, image and video cost $0.30.
- Output price includes thinking tokens.
- Gemini 3.8 Flash
- Introductory price through December 31, 2026. From January 1, 2027: $1.50 input, $0.15 cached, $7.50 output per 1M tokens.
- Output price includes thinking tokens.
- Priority inference costs 1.8x the standard rate.
- Gemini 3.5 Flash
- Output price includes thinking tokens.
- Context cache storage costs $1.00 per 1M tokens per hour.
- Priority inference costs 1.8x the standard rate.
- Gemini 2.5 Pro
- Prompts over 200K tokens are billed at $2.50 input / $15 output per 1M tokens.
- Output price includes thinking tokens.
- Gemini 3.1 Pro
- Prompts over 200K tokens are billed at $4 input / $18 output per 1M tokens.
- No free tier on the Gemini API.
- Output price includes thinking tokens.
Recent Gemini pricing changes
Gemini 3.8 Flash introductory pricing ends
The standard price rises from $0.75 to $1.50 per 1M input tokens, from $0.075 to $0.15 per 1M cached input tokens and from $3.75 to $7.50 per 1M output tokens.
Google’s Gemini 3.8 text-to-speech prices double
Gemini 3.8 Flash TTS goes from $0.50 to $1.00 per 1M text input tokens and from $9.00 to $18.00 per 1M audio output tokens (about $0.0135 to $0.027 per minute of audio). Gemini 3.8 Flash-Lite TTS goes from $6.00 to $12.00 per 1M audio output tokens. Batch, Flex and Priority rates double too.
Gemini against other providers
Subscription instead of the API?
These plans run Gemini models for a flat monthly price. See whether one costs less than paying per token with the subscription vs API calculator.
Frequently asked questions
- How much does the Gemini API cost?
- Gemini API prices range from $0.30 input / $2.50 output per 1M tokens for Gemini 3.5 Flash-Lite to $2.00 / $12.00 for Gemini 3.1 Pro. A support chatbot handling 100,000 replies a month costs about $145.00 on Gemini 3.5 Flash-Lite.
- What is the cheapest Gemini model?
- Gemini 3.5 Flash-Lite, at a blended $0.85 per 1M tokens (3 input : 1 output) — 81% less per token than Gemini 3.1 Pro.
- Does the Gemini API have a batch discount?
- Yes. Batch requests, which return results within a window of up to 24 hours instead of immediately, cost 50% less on the Gemini models that support it.
- Which Gemini model is the most capable?
- Gemini 3.8 Flash has the highest Epoch Capabilities Index of the Gemini models we track (156.5), at $0.75 input / $3.75 output per 1M tokens.