Google Gemini API pricing
Gemini API rates and request cost estimates for text and multimodal workloads.
Compare one workload
Rates only become comparable when input, output and request count are identical. The calculation below uses the same workload for every model.
- Input per request
- 2,000
- Output per request
- 1,000
- Requests
- 1,000
Current model costs
| Model | Input / 1M | Output / 1M | Workload estimate |
|---|---|---|---|
| Google: Gemini 3.8 FlashGoogle | $0.75 | $3.75 | $5.25 |
| Google: Gemini 3.7 FlashGoogle | $0.75 | $3.75 | $5.25 |
| Google: Gemini 3.6 FlashGoogle | $0.75 | $3.75 | $5.25 |
| Google: Gemini 3.5 Flash LiteGoogle | $0.3 | $2.5 | $3.1 |
| Google: Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)Google | $0.25 | $1.5 | $2 |
| Google: Nano Banana 2 (Gemini 3.1 Flash Image)Google | $0.5 | $3 | $4 |
| Google: Nano Banana Pro (Gemini 3 Pro Image)Google | $2 | $12 | $16 |
| Google: Gemini 3.5 FlashGoogle | $1.5 | $9 | $12 |
| Google: Gemini 3.1 Flash LiteGoogle | $0.25 | $1.5 | $2 |
| Google: Gemma 4 26B A4B Google | $0.08 | $0.26 | $0.41 |
Estimate uses published token rates. Caching, reasoning tokens, tools, images, routing and taxes can change the final bill.
How much does the Gemini API cost?
Start with the work unit you can measure: one conversation, document, article or agent run. Record both prompt and response length, then multiply by real monthly volume. Output often costs more than input, so a concise answer can matter more than a short prompt.
Price is only one constraint. Check context length, modalities, tool support, latency and output quality before choosing a production model. Test a small representative sample instead of relying on a single benchmark.