Grok 4.3 vs DeepSeek V4 Flash

DeepSeek V4 Flash is the cheaper of the two across every workload shape tested below. Headline rates are $1.25 in / $2.5 out for Grok 4.3 versus $0.44 in / $1.32 out for DeepSeek V4 Flash, per million tokens.

Real monthly cost, three workloads

Workload Grok 4.3 DeepSeek V4 Flash Cheaper
Support chatbot
1.5k in / 400 out · 100k req/mo · 30% cached
$240.25 $99.63 DeepSeek V4 Flash (2.4× cheaper)
RAG / doc Q&A
12k in / 700 out · 50k req/mo · 60% cached
$459.50 $156.84 DeepSeek V4 Flash (2.9× cheaper)
Coding agent
25k in / 2.5k out · 20k req/mo · 80% cached
$330.00 $115.60 DeepSeek V4 Flash (2.9× cheaper)

Costs include prompt caching at the stated hit rate and per-model tokenizer correction. Adjust for your own traffic →

Specifications side by side

Grok 4.3 DeepSeek V4 Flash
Provider xAI DeepSeek
Input / M tokens $1.25 $0.44
Output / M tokens $2.5 $1.32
Cached input $0.2 $0.014
Context window 1000K 1000K
Batch discount — —
Tokenizer factor ×1 ×1
Min prefix to cache — —
Announced retirement — —

Which should you pick?

On cost alone, DeepSeek V4 Flash wins every workload shape above.

Full detail: Grok 4.3 · DeepSeek V4 Flash

All figures verified 2026-08-18 against official provider documentation. Fields a provider does not publish are shown as “—” rather than estimated.