Claude Haiku 4.5 vs DeepSeek V4 Pro

DeepSeek V4 Pro is the cheaper of the two across most of the workload shapes tested below. Headline rates are $1 in / $5 out for Claude Haiku 4.5 versus $1.32 in / $3.96 out for DeepSeek V4 Pro, per million tokens.

Real monthly cost, three workloads

Workload Claude Haiku 4.5 DeepSeek V4 Pro Cheaper
Support chatbot
1.5k in / 400 out · 100k req/mo · 30% cached
$309.50 $298.98 DeepSeek V4 Pro (1.0× cheaper)
RAG / doc Q&A
12k in / 700 out · 50k req/mo · 60% cached
$451.00 $471.24 Claude Haiku 4.5 (1.0× cheaper)
Coding agent
25k in / 2.5k out · 20k req/mo · 80% cached
$390.00 $347.60 DeepSeek V4 Pro (1.1× cheaper)

Costs include prompt caching at the stated hit rate and per-model tokenizer correction. Adjust for your own traffic →

Specifications side by side

Claude Haiku 4.5 DeepSeek V4 Pro
Provider Anthropic DeepSeek
Input / M tokens $1 $1.32
Output / M tokens $5 $3.96
Cached input $0.1 $0.044
Context window 200K 1000K
Batch discount −50% —
Tokenizer factor ×1 ×1
Min prefix to cache 4096 —
Announced retirement 2026-10-15+ —

Which should you pick?

On cost alone, DeepSeek V4 Pro wins the majority of the shapes above. The ranking flips depending on the input-to-output ratio, so check the shape closest to your own traffic.

Caching behaviour differs. Claude Haiku 4.5 requires a prefix of at least 4,096 tokens before caching activates. Below those thresholds cache_control is ignored silently, with no error — which can quietly erase the saving you were counting on.

Lifecycle matters as much as price here. Claude Haiku 4.5 has an announced shutdown date of 2026-10-15 or later. A model that is marginally cheaper but retires within your planning horizon is rarely the better choice.

Full detail: Claude Haiku 4.5 · DeepSeek V4 Pro

All figures verified 2026-08-18 against official provider documentation. Fields a provider does not publish are shown as “—” rather than estimated.