Prices checked 2026-10-02

Gemini API Pricing (October 2026)

Gemini API pricing runs from $0.30 per 1M input tokens on Gemini 3.5 Flash-Lite to $12 per 1M output tokens on Gemini 3.1 Pro. Gemini 3.1 Pro costs $2 input and $12 output per 1M tokens. Output tokens cost 5-8x more than input, and batch processing halves both.

Current Gemini models

Prices per 1 million tokens. Cached input is the price of reading a cached prompt prefix.

ModelInput / 1MCached inputOutput / 1MBatch in / outContext
Gemini 3.1 Pro$2$0.20$12$1 / $61M
Gemini 3.8 Flash$0.75$0.075$3.75$0.375 / $1.8751M
Gemini 3.5 Flash-Lite$0.30$0.03$2.50$0.15 / $1.251M

What things actually cost

Four common jobs, priced on every current Gemini model.

JobGemini 3.5 Flash-LiteGemini 3.8 FlashGemini 3.1 Pro
Summarize a 2,000-word article2,667 input tokens in, a 300-word summary (400 tokens) out; total for 1,000 articles$1.80$3.50$10.13
Support chatbot, 10,000 messages a dayeach reply reads ~1,000 tokens of conversation and writes ~200 tokens; total for 30 days$240$450$1,320
Review a 5,000-line codebase50,000 input tokens at ~10 tokens per line, a 1,500-word review out; total for one review$0.02$0.04$0.12
Read a 300-page book and answer questions200,100 input tokens at ~667 tokens per page, 10 answers of ~200 words; total for one book$0.07$0.16$0.43

Standard (non-batch, uncached) prices. Token counts use 1 token ≈ 0.75 English words; your real counts depend on each model's tokenizer.

Calculate your own budget

Pick a model and a budget to see how many tokens, words and pages it buys.

Token direction

833,333tokens

625,000English words
6.9novels
1,249pages of text
83,333lines of code
833images analyzed
69.4hours of speech

Discounts and pricing rules

  • Context caching costs 10% of the input price.
  • The Batch API costs half the standard rate on input and output.
  • Gemini 3.1 Pro costs $4 input / $18 output per 1M tokens for prompts above 200K tokens (context caching $0.40, batch $2 / $9).
  • Gemini 3.8 Flash is at an introductory $0.75 / $3.75 through December 31, 2026, rising to $1.50 / $7.50 on January 1, 2027.
  • Prices shown are paid-tier text prices. Audio input is priced separately on some models.

Source: Google's official pricing page. Base prices are re-checked daily against live market data.

Previous-generation Gemini models

Still available through the API. Each row names its current replacement.

ModelInput / 1MCached inputOutput / 1MBatch in / outContextReplaced by
Gemini 3.5 Flash$1.50$0.15$9$0.75 / $4.501MGemini 3.8 Flash
Gemini 3 Flash$0.50$0.05$3$0.25 / $1.501MGemini 3.8 Flash

Frequently asked questions

How much does the Gemini API cost?
Gemini API prices run from $0.30 per 1M input tokens to $12 per 1M output tokens, depending on the model. Current models (input / output per 1M tokens): Gemini 3.1 Pro $2 / $12, Gemini 3.8 Flash $0.75 / $3.75, Gemini 3.5 Flash-Lite $0.30 / $2.50.
What is the cheapest Gemini model?
Gemini 3.5 Flash-Lite at $0.30 per 1M input tokens and $2.50 per 1M output tokens, or $0.15 / $1.25 through the Batch API. One million input tokens is about 750,000 words.
How much does Gemini 3.1 Pro cost?
Gemini 3.1 Pro costs $2 per 1M input tokens and $12 per 1M output tokens, with a 1M-token context window. It is the most expensive current Gemini model.
How can I lower my Gemini API bill?
Context caching costs 10% of the input price. The Batch API costs half the standard rate on input and output.
Is the Gemini API the same as a subscription?
Gemini app subscriptions (Google AI plans) are billed separately from the Gemini API, which is pay-per-token.