Gemini 3.7 Flash pricing

Gemini 3.7 Flash costs $0.75 per million input tokens and $3.75 per million output tokens , with cached input at $0.075. Batch requests are discounted 50%.

Price per million tokens

Input $0.75
Cached input $0.075
Output $3.75
Context window not published
Batch discount −50%

This rate is valid through 2026-12-31 and doubles on 2027-01-01

What it costs in practice

Monthly cost for three common workload shapes, with prompt caching applied.

Workload Shape Per request Per month
Support chatbot 1.5k in / 400 out, 100k requests/month, 30% cached $0.00232 $232.13
RAG / document Q&A 12k in / 700 out, 50k requests/month, 60% cached $0.00677 $338.25
Coding agent 25k in / 2.5k out, 20k requests/month, 80% cached $0.0146 $292.50

Model your own workload with caching, batching and tokenizer correction →

Comparable models from other providers

Model Provider Input Output
GPT-5.4 mini OpenAI $0.75 $4.5
Claude Haiku 4.5 Anthropic $1 $5
Grok Build 0.1 xAI $1 $2
o4-mini OpenAI $1.1 $4.4

All figures verified 2026-08-18 against Google's official pricing documentation . Fields the provider does not publish are shown as “—” rather than estimated.