Gemini 3.1 Pro (preview) pricing

Gemini 3.1 Pro (preview) costs $2 per million input tokens and $12 per million output tokens , with cached input at $0.2. Batch requests are discounted 50%.

Price per million tokens

Input $2
Cached input $0.2
Output $12
Context window not published
Batch discount −50%

This model uses tiered pricing: above 200,000 input tokens the rate rises to $4 in / $18 out.

What it costs in practice

Monthly cost for three common workload shapes, with prompt caching applied.

Workload Shape Per request Per month
Support chatbot 1.5k in / 400 out, 100k requests/month, 30% cached $0.00699 $699.00
RAG / document Q&A 12k in / 700 out, 50k requests/month, 60% cached $0.0194 $972.00
Coding agent 25k in / 2.5k out, 20k requests/month, 80% cached $0.044 $880.00

Model your own workload with caching, batching and tokenizer correction →

Comparable models from other providers

Model Provider Input Output
Claude Sonnet 5 Anthropic $2 $10
GPT-5.6 Terra OpenAI $2 $12
GPT-4.1 OpenAI $2 $8
o3 OpenAI $2 $8

All figures verified 2026-08-18 against Google's official pricing documentation . Fields the provider does not publish are shown as “—” rather than estimated.