Gemini 3.7 Flash vs Grok 4.3

Gemini 3.7 Flash is the cheaper of the two across every workload shape tested below. Headline rates are $0.75 in / $3.75 out for Gemini 3.7 Flash versus $1.25 in / $2.5 out for Grok 4.3, per million tokens.

Real monthly cost, three workloads

Workload Gemini 3.7 Flash Grok 4.3 Cheaper
Support chatbot
1.5k in / 400 out · 100k req/mo · 30% cached
$232.13 $240.25 Gemini 3.7 Flash (1.0× cheaper)
RAG / doc Q&A
12k in / 700 out · 50k req/mo · 60% cached
$338.25 $459.50 Gemini 3.7 Flash (1.4× cheaper)
Coding agent
25k in / 2.5k out · 20k req/mo · 80% cached
$292.50 $330.00 Gemini 3.7 Flash (1.1× cheaper)

Costs include prompt caching at the stated hit rate and per-model tokenizer correction. Adjust for your own traffic →

Specifications side by side

Gemini 3.7 Flash Grok 4.3
Provider Google xAI
Input / M tokens $0.75 $1.25
Output / M tokens $3.75 $2.5
Cached input $0.075 $0.2
Context window — 1000K
Batch discount −50% —
Tokenizer factor ×1 ×1
Min prefix to cache — —
Announced retirement — —

Which should you pick?

On cost alone, Gemini 3.7 Flash wins every workload shape above.

Full detail: Gemini 3.7 Flash · Grok 4.3

All figures verified 2026-08-18 against official provider documentation. Fields a provider does not publish are shown as “—” rather than estimated.