Gemini 2.5 Flash pricing

Gemini 2.5 Flash costs $0.3 per million input tokens and $2.5 per million output tokens , with cached input at $0.03. Its context window is 1M tokens. Batch requests are discounted 50%.

Price per million tokens

Input $0.3
Cached input $0.03
Output $2.5
Context window 1M tokens
Batch discount −50%

What it costs in practice

Monthly cost for three common workload shapes, with prompt caching applied.

Workload Shape Per request Per month
Support chatbot 1.5k in / 400 out, 100k requests/month, 30% cached $0.00133 $132.85
RAG / document Q&A 12k in / 700 out, 50k requests/month, 60% cached $0.00341 $170.30
Coding agent 25k in / 2.5k out, 20k requests/month, 80% cached $0.00835 $167.00

Model your own workload with caching, batching and tokenizer correction →

Comparable models from other providers

Model Provider Input Output
GPT-5 mini OpenAI $0.25 $2
GPT-4.1 mini OpenAI $0.4 $1.6
DeepSeek V4 Flash DeepSeek $0.44 $1.32
GPT-5.6 Luna OpenAI $0.2 $1.2

All figures verified 2026-08-18 against Google's official pricing documentation . Fields the provider does not publish are shown as “—” rather than estimated.