Provider
Google API pricing: every model compared
Google offers 20 models through its API in our data, priced from $0.05 to $2 per million input tokens. The newest is Gemini 3.8 Flash, first listed on 2026-09-02.
- Models
- 20
- Cheapest input
- $0.05
- Priciest input
- $2
- Largest context
- 1,048,576
- With caching
- 17 of 20
- Open weights
- 6 of 20
Cheapest
Gemma 3 4BGoogle$0.05 in · $0.10 outLargest context
Gemini 2.5 FlashGoogle1,048,576 tokensLine-up
Every Google model and price
| Model | Input / 1M | Output / 1M | Context | Released |
|---|---|---|---|---|
| Gemini 3.8 Flash | $0.75 | $3.75 | 1,048,576 | 2026-09-02 |
| Gemini 3.7 Flash | $0.75 | $3.75 | 1,048,576 | 2026-08-13 |
| Gemini 3.5 Flash Lite | $0.30 | $2.50 | 1,048,576 | 2026-07-21 |
| Gemini 3.6 Flash | $0.75 | $3.75 | 1,048,576 | 2026-07-21 |
| Gemini 3.5 Flash | $1.50 | $9 | 1,048,576 | 2026-05-19 |
| Gemini 3.1 Flash Lite | $0.25 | $1.50 | 1,048,576 | 2026-05-07 |
| Gemma 4 26B A4B | $0.0765 | $0.255 | 262,144 | 2026-04-03 |
| Gemma 4 31B | $0.09 | $0.34 | 262,144 | 2026-04-02 |
| Gemini 3.1 Flash Lite Preview | $0.25 | $1.50 | 1,048,576 | 2026-03-03 |
| Gemini 3.1 Pro Preview Custom Tools | $2 | $12 | 1,048,576 | 2026-02-25 |
| Gemini 3.1 Pro Preview | $2 | $12 | 1,048,576 | 2026-02-19 |
| Gemini 3 Flash Preview | $0.50 | $3 | 1,048,576 | 2025-12-17 |
| Gemini 2.5 Flash Lite | $0.10 | $0.40 | 1,048,576 | 2025-07-22 |
| Gemini 2.5 Flash | $0.30 | $2.50 | 1,048,576 | 2025-06-17 |
| Gemini 2.5 Pro | $1.25 | $10 | 1,048,576 | 2025-06-17 |
| Gemini 2.5 Pro Preview 06-05 | $1.25 | $10 | 1,048,576 | 2025-06-05 |
| Gemma 3 12B | $0.05 | $0.15 | 131,072 | 2025-03-13 |
| Gemma 3 4B | $0.05 | $0.10 | 131,072 | 2025-03-13 |
| Gemma 3 27B | $0.08 | $0.45 | 131,072 | 2025-03-12 |
| Gemma 2 27B | $0.65 | $0.65 | 8,192 | 2024-07-13 |
Prices in US dollars per 1M tokens, newest first. Release dates are when our source first listed the model.
Overview
Google’s pricing at a glance
The spread between Google’s cheapest and most expensive model is large: Gemini 3.1 Pro Preview Custom Tools costs 40× more per input token than Gemma 3 4B. Picking the smallest model that does the job well is usually the biggest saving available, ahead of any discount.
17 of 20 of the line-up has a cached-input price, which helps chatbots and agents that resend the same instructions. 12 of 20 can run through a batch API at a discount, for work that can wait up to a day. 19 of 20 accept images as input.
FAQ
Google API questions
How much does the Google API cost?
Google's 20 models range from $0.05 to $2 per million input tokens, and output costs more than input on almost every model. The exact bill depends on your token volumes, which you can estimate in the LLM cost calculator.
What is Google's cheapest model?
By blended price (three parts input to one part output) it is Gemma 3 4B, at $0.05 input and $0.10 output per million tokens. The most expensive is Gemini 3.1 Pro Preview Custom Tools.
Which Google model has the largest context window?
Gemini 2.5 Flash, with 1,048,576 tokens per request.
Does Google offer prompt caching or batch discounts?
In our data, 17 of 20 of Google's models have a published cached-input price and 12 of 20 have a published batch price. Both can cut costs substantially for repeated prompts or work that can wait.