Gemini 3.5 Flash API pricing
Google model · context 1,048,576 tokens · max output 65,536 · vision, reasoning. Last price check 2026-10-07. Official pricing.
| Input | $1.50 / 1M tokens |
|---|---|
| Output | $9.00 / 1M tokens |
| Cached input | $0.15 / 1M tokens |
| Batch input | $0.75 / 1M tokens |
| Batch output | $4.50 / 1M tokens |
What it costs per month
| Scenario (per month) | Gemini 3.5 Flash |
|---|---|
| Customer support chatbot 1,000 req/day · 1,500 in · 300 out | $118 |
| RAG over documents 500 req/day · 6,000 in · 400 out | $153 |
| Agent loop (10 calls/task, 200 tasks) 2,000 req/day · 8,000 in · 500 out | $536 |
| Batch summarization 10,000 req/day · 3,000 in · 200 out · batch | $945 |
Adjust the numbers in the calculator.
Cheaper alternatives
Lowest monthly cost for the chatbot scenario among reasoning models:
- GPT-5 nano — $4.84/mo
- Gemini 2.5 Flash Lite — $6.08/mo
- GPT-6 Luna — $6.97/mo
- Magistral Small (latest) — $9.11/mo
- Mistral Small (latest) — $9.11/mo
Price history
No price changes recorded since tracking began on 2026-10-07.