# Gemini API pricing, every Google model per million tokens

On 25 September 2026, Google models on this list cost from $0.05 per million input tokens (Gemma 3 4B) to $2.00 (Nano Banana Pro (Gemini 3 Pro Image)), and from $0.10 per million output tokens (Gemma 3 4B) to $12.00 (Nano Banana Pro (Gemini 3 Pro Image)).

Prices checked 25 September 2026 at 06.00 UTC . Source OpenRouter public model list . US dollars per million tokens

Input tokens per call

Output tokens per call

Calls a day

Search models

| Model | Input | Output | Cached input | Batch in / out | Context | Your month | Listed |
|---|---|---|---|---|---|---|---|
| Gemini 3.8 Flash | $0.75 | $3.75 | $0.075 | $0.375 / $1.88 | 1.04858M | $101.25 | 2026-09-02 |
| Gemini 3.7 Flash | $0.75 | $3.75 | $0.075 | $0.375 / $1.88 | 1.04858M | $101.25 | 2026-08-13 |
| Gemini 3.6 Flash | $0.75 | $3.75 | $0.075 | $0.375 / $1.88 | 1.04858M | $101.25 | 2026-07-21 |
| Gemini 3.5 Flash Lite | $0.30 | $2.50 | $0.03 | $0.15 / $1.25 | 1.04858M | $55.50 | 2026-07-21 |
| Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) | $0.25 | $1.50 |  |  | 66K | $37.50 | 2026-06-30 |
| Nano Banana 2 (Gemini 3.1 Flash Image) | $0.50 | $3.00 |  |  | 131K | $75.00 | 2026-06-18 |
| Nano Banana Pro (Gemini 3 Pro Image) | $2.00 | $12.00 | $0.20 |  | 131K | $300.00 | 2026-06-18 |
| Gemini 3.5 Flash | $1.50 | $9.00 | $0.15 | $0.75 / $4.50 | 1.04858M | $225.00 | 2026-05-19 |
| Gemini 3.1 Flash Lite | $0.25 | $1.50 | $0.025 | $0.125 / $0.75 | 1.04858M | $37.50 | 2026-05-07 |
| Gemma 4 26B A4B free tier on OpenRouter | $0.0675 | $0.225 | $0.0375 |  | 262K | $7.42 | 2026-04-03 |
| Gemma 4 31B free tier on OpenRouter | $0.09 | $0.34 | $0.05 |  | 262K | $10.50 | 2026-04-02 |
| Gemini 3.1 Flash Lite Preview | $0.25 | $1.50 | $0.025 |  | 1.04858M | $37.50 | 2026-03-03 |
| Nano Banana 2 (Gemini 3.1 Flash Image Preview) | $0.50 | $3.00 |  |  | 66K | $75.00 | 2026-02-26 |
| Gemini 3.1 Pro Preview Custom Tools | $2.00 | $12.00 | $0.20 |  | 1.04858M | $300.00 | 2026-02-25 |
| Gemini 3.1 Pro Preview | $2.00 | $12.00 | $0.20 | $1.00 / $6.00 | 1.04858M | $300.00 | 2026-02-19 |
| Gemini 3 Flash Preview | $0.50 | $3.00 | $0.05 | $0.25 / $1.50 | 1.04858M | $75.00 | 2025-12-17 |
| Nano Banana Pro (Gemini 3 Pro Image Preview) | $2.00 | $12.00 | $0.20 |  | 66K | $300.00 | 2025-11-20 |
| Nano Banana (Gemini 2.5 Flash Image) | $0.30 | $2.50 | $0.03 |  | 33K | $55.50 | 2025-10-07 |
| Gemini 2.5 Flash Lite retiring 2026-10-20 | $0.10 | $0.40 | $0.01 | $0.05 / $0.20 | 1.04858M | $12.00 | 2025-07-22 |
| Gemini 2.5 Flash retiring 2026-10-20 | $0.30 | $2.50 | $0.03 | $0.15 / $1.25 | 1.04858M | $55.50 | 2025-06-17 |
| Gemini 2.5 Pro retiring 2026-10-20 | $1.25 | $10.00 | $0.125 | $0.625 / $5.00 | 1.04858M | $225.00 | 2025-06-17 |
| Gemini 2.5 Pro Preview 06-05 | $1.25 | $10.00 | $0.125 |  | 1.04858M | $225.00 | 2025-06-05 |
| Gemma 3 4B | $0.05 | $0.10 |  |  | 131K | $4.50 | 2025-03-13 |
| Gemma 3 12B | $0.05 | $0.15 |  |  | 131K | $5.25 | 2025-03-13 |
| Gemma 3 27B | $0.08 | $0.45 | $0.04 |  | 131K | $11.55 | 2025-03-12 |
| Gemma 2 27B | $0.65 | $0.65 |  |  | 8K | $48.75 | 2024-07-13 |

[Compare Google with every other provider](https://shurco.ai/tools/llm-api-pricing/)

## Common questions

### How much does the Gemini API cost?

On 25 September 2026, Google models on this list cost from $0.05 per million input tokens (Gemma 3 4B) to $2.00 (Nano Banana Pro (Gemini 3 Pro Image)), and from $0.10 per million output tokens (Gemma 3 4B) to $12.00 (Nano Banana Pro (Gemini 3 Pro Image)).

### What is the cheapest Gemini model?

For a typical mix of three input tokens to every output token, the cheapest is Gemma 3 4B at $0.05 per million input tokens and $0.10 per million output tokens.

### What is the newest Gemini model on the list?

Gemini 3.8 Flash, listed on 2 September 2026, at $0.75 per million input tokens and $3.75 per million output tokens.

### Which Gemini model has the largest context window?

Gemini 3.8 Flash, with 1.04858M tokens of context.

### Does Google offer batch pricing?

Yes, 11 of its models on this list show a batch price, with batch input at a median of 50 percent of the standard price. Batch suits work that can wait for its answer.

## Other providers

[GPT](https://shurco.ai/tools/llm-api-pricing/openai/)[Claude](https://shurco.ai/tools/llm-api-pricing/anthropic/)[DeepSeek](https://shurco.ai/tools/llm-api-pricing/deepseek/)[Mistral](https://shurco.ai/tools/llm-api-pricing/mistral/)[Llama](https://shurco.ai/tools/llm-api-pricing/meta/)[Grok](https://shurco.ai/tools/llm-api-pricing/xai/)[Qwen](https://shurco.ai/tools/llm-api-pricing/qwen/)[Kimi](https://shurco.ai/tools/llm-api-pricing/moonshot/)[GLM](https://shurco.ai/tools/llm-api-pricing/zai/)[Command](https://shurco.ai/tools/llm-api-pricing/cohere/)[Nova](https://shurco.ai/tools/llm-api-pricing/amazon/)[MiniMax](https://shurco.ai/tools/llm-api-pricing/minimax/)

## Choosing a model for a production agent

Price is one input. Quality on your own cases, latency and the cost of retries decide the rest. [Tell me what you are building.](https://shurco.ai/#contact)

Web version https://shurco.ai/tools/llm-api-pricing/google/
