Gemini 3.5 Flash pricing
Gemini 3.5 Flash costs $1.5 per million input tokens and $9 per million output tokens , with cached input at $0.15. Batch requests are discounted 50%.
Price per million tokens
| Input | $1.5 |
| Cached input | $0.15 |
| Output | $9 |
| Context window | not published |
| Batch discount | −50% |
What it costs in practice
Monthly cost for three common workload shapes, with prompt caching applied.
| Workload | Shape | Per request | Per month |
|---|---|---|---|
| Support chatbot | 1.5k in / 400 out, 100k requests/month, 30% cached | $0.00524 | $524.25 |
| RAG / document Q&A | 12k in / 700 out, 50k requests/month, 60% cached | $0.0146 | $729.00 |
| Coding agent | 25k in / 2.5k out, 20k requests/month, 80% cached | $0.033 | $660.00 |
Model your own workload with caching, batching and tokenizer correction →
Comparable models from other providers
| Model | Provider | Input | Output |
|---|---|---|---|
| DeepSeek V4 Pro | DeepSeek | $1.32 | $3.96 |
| GPT-5.3 Codex | OpenAI | $1.75 | $14 |
| GPT-5.2 | OpenAI | $1.75 | $14 |
| GPT-5.1 | OpenAI | $1.25 | $10 |
All figures verified 2026-08-18 against Google's official pricing documentation . Fields the provider does not publish are shown as “—” rather than estimated.