GPT-5 nano vs Gemini 2.5 Flash-Lite

GPT-5 nano is the cheaper of the two across every workload shape tested below. Headline rates are $0.05 in / $0.4 out for GPT-5 nano versus $0.1 in / $0.4 out for Gemini 2.5 Flash-Lite, per million tokens.

Real monthly cost, three workloads

Workload GPT-5 nano Gemini 2.5 Flash-Lite Cheaper
Support chatbot
1.5k in / 400 out · 100k req/mo · 30% cached
$21.48 $26.95 GPT-5 nano (1.3× cheaper)
RAG / doc Q&A
12k in / 700 out · 50k req/mo · 60% cached
$27.80 $41.60 GPT-5 nano (1.5× cheaper)
Coding agent
25k in / 2.5k out · 20k req/mo · 80% cached
$27.00 $34.00 GPT-5 nano (1.3× cheaper)

Costs include prompt caching at the stated hit rate and per-model tokenizer correction. Adjust for your own traffic →

Specifications side by side

GPT-5 nano Gemini 2.5 Flash-Lite
Provider OpenAI Google
Input / M tokens $0.05 $0.1
Output / M tokens $0.4 $0.4
Cached input $0.005 $0.01
Context window — —
Batch discount −50% −50%
Tokenizer factor ×1 ×1
Min prefix to cache — —
Announced retirement — —

Which should you pick?

On cost alone, GPT-5 nano wins every workload shape above.

Full detail: GPT-5 nano · Gemini 2.5 Flash-Lite

All figures verified 2026-08-18 against official provider documentation. Fields a provider does not publish are shown as “—” rather than estimated.