Claude Haiku 4.5 vs Grok 4.3
Grok 4.3 is the cheaper of the two across most of the workload shapes tested below. Headline rates are $1 in / $5 out for Claude Haiku 4.5 versus $1.25 in / $2.5 out for Grok 4.3, per million tokens.
Real monthly cost, three workloads
| Workload | Claude Haiku 4.5 | Grok 4.3 | Cheaper |
|---|---|---|---|
| Support chatbot 1.5k in / 400 out · 100k req/mo · 30% cached | $309.50 | $240.25 | Grok 4.3 (1.3× cheaper) |
| RAG / doc Q&A 12k in / 700 out · 50k req/mo · 60% cached | $451.00 | $459.50 | Claude Haiku 4.5 (1.0× cheaper) |
| Coding agent 25k in / 2.5k out · 20k req/mo · 80% cached | $390.00 | $330.00 | Grok 4.3 (1.2× cheaper) |
Costs include prompt caching at the stated hit rate and per-model tokenizer correction. Adjust for your own traffic →
Specifications side by side
| Claude Haiku 4.5 | Grok 4.3 | |
|---|---|---|
| Provider | Anthropic | xAI |
| Input / M tokens | $1 | $1.25 |
| Output / M tokens | $5 | $2.5 |
| Cached input | $0.1 | $0.2 |
| Context window | 200K | 1000K |
| Batch discount | −50% | — |
| Tokenizer factor | ×1 | ×1 |
| Min prefix to cache | 4096 | — |
| Announced retirement | 2026-10-15+ | — |
Which should you pick?
On cost alone, Grok 4.3 wins the majority of the shapes above. The ranking flips depending on the input-to-output ratio, so check the shape closest to your own traffic.
Caching behaviour differs. Claude Haiku 4.5 requires a prefix of at least 4,096 tokens before caching activates.
Below those thresholds cache_control is ignored silently, with no error — which can quietly
erase the saving you were counting on.
Lifecycle matters as much as price here. Claude Haiku 4.5 has an announced shutdown date of 2026-10-15 or later. A model that is marginally cheaper but retires within your planning horizon is rarely the better choice.
Full detail: Claude Haiku 4.5 · Grok 4.3
All figures verified 2026-08-18 against official provider documentation. Fields a provider does not publish are shown as “—” rather than estimated.