Prices checked 2026-10-02
Gemini API Pricing (October 2026)
Gemini API pricing runs from $0.30 per 1M input tokens on Gemini 3.5 Flash-Lite to $12 per 1M output tokens on Gemini 3.1 Pro. Gemini 3.1 Pro costs $2 input and $12 output per 1M tokens. Output tokens cost 5-8x more than input, and batch processing halves both.
Current Gemini models
Prices per 1 million tokens. Cached input is the price of reading a cached prompt prefix.
| Model | Input / 1M | Cached input | Output / 1M | Batch in / out | Context |
|---|---|---|---|---|---|
| Gemini 3.1 Pro | $2 | $0.20 | $12 | $1 / $6 | 1M |
| Gemini 3.8 Flash | $0.75 | $0.075 | $3.75 | $0.375 / $1.875 | 1M |
| Gemini 3.5 Flash-Lite | $0.30 | $0.03 | $2.50 | $0.15 / $1.25 | 1M |
What things actually cost
Four common jobs, priced on every current Gemini model.
| Job | Gemini 3.5 Flash-Lite | Gemini 3.8 Flash | Gemini 3.1 Pro |
|---|---|---|---|
| Summarize a 2,000-word article2,667 input tokens in, a 300-word summary (400 tokens) out; total for 1,000 articles | $1.80 | $3.50 | $10.13 |
| Support chatbot, 10,000 messages a dayeach reply reads ~1,000 tokens of conversation and writes ~200 tokens; total for 30 days | $240 | $450 | $1,320 |
| Review a 5,000-line codebase50,000 input tokens at ~10 tokens per line, a 1,500-word review out; total for one review | $0.02 | $0.04 | $0.12 |
| Read a 300-page book and answer questions200,100 input tokens at ~667 tokens per page, 10 answers of ~200 words; total for one book | $0.07 | $0.16 | $0.43 |
Standard (non-batch, uncached) prices. Token counts use 1 token ≈ 0.75 English words; your real counts depend on each model's tokenizer.
Calculate your own budget
Pick a model and a budget to see how many tokens, words and pages it buys.
833,333tokens
Discounts and pricing rules
- Context caching costs 10% of the input price.
- The Batch API costs half the standard rate on input and output.
- Gemini 3.1 Pro costs $4 input / $18 output per 1M tokens for prompts above 200K tokens (context caching $0.40, batch $2 / $9).
- Gemini 3.8 Flash is at an introductory $0.75 / $3.75 through December 31, 2026, rising to $1.50 / $7.50 on January 1, 2027.
- Prices shown are paid-tier text prices. Audio input is priced separately on some models.
Source: Google's official pricing page. Base prices are re-checked daily against live market data.
Previous-generation Gemini models
Still available through the API. Each row names its current replacement.
| Model | Input / 1M | Cached input | Output / 1M | Batch in / out | Context | Replaced by |
|---|---|---|---|---|---|---|
| Gemini 3.5 Flash | $1.50 | $0.15 | $9 | $0.75 / $4.50 | 1M | Gemini 3.8 Flash |
| Gemini 3 Flash | $0.50 | $0.05 | $3 | $0.25 / $1.50 | 1M | Gemini 3.8 Flash |
Frequently asked questions
- How much does the Gemini API cost?
- Gemini API prices run from $0.30 per 1M input tokens to $12 per 1M output tokens, depending on the model. Current models (input / output per 1M tokens): Gemini 3.1 Pro $2 / $12, Gemini 3.8 Flash $0.75 / $3.75, Gemini 3.5 Flash-Lite $0.30 / $2.50.
- What is the cheapest Gemini model?
- Gemini 3.5 Flash-Lite at $0.30 per 1M input tokens and $2.50 per 1M output tokens, or $0.15 / $1.25 through the Batch API. One million input tokens is about 750,000 words.
- How much does Gemini 3.1 Pro cost?
- Gemini 3.1 Pro costs $2 per 1M input tokens and $12 per 1M output tokens, with a 1M-token context window. It is the most expensive current Gemini model.
- How can I lower my Gemini API bill?
- Context caching costs 10% of the input price. The Batch API costs half the standard rate on input and output.
- Is the Gemini API the same as a subscription?
- Gemini app subscriptions (Google AI plans) are billed separately from the Gemini API, which is pay-per-token.