Claude Opus 4.6 pricing
Claude Opus 4.6 costs $5 per million input tokens and $25 per million output tokens , with cached input at $0.5. Its context window is 1M tokens. Batch requests are discounted 50%.
claude-opus-5. See all shutdown dates → Price per million tokens
| Input | $5 |
| Cached input | $0.5 |
| Output | $25 |
| Context window | 1M tokens |
| Batch discount | −50% |
| Minimum prefix to cache | 4,096 tokens |
What it costs in practice
Monthly cost for three common workload shapes, with prompt caching applied.
| Workload | Shape | Per request | Per month |
|---|---|---|---|
| Support chatbot | 1.5k in / 400 out, 100k requests/month, 30% cached | $0.0155 | $1,548 |
| RAG / document Q&A | 12k in / 700 out, 50k requests/month, 60% cached | $0.0451 | $2,255 |
| Coding agent | 25k in / 2.5k out, 20k requests/month, 80% cached | $0.0975 | $1,950 |
Model your own workload with caching, batching and tokenizer correction →
Caching on Claude Opus 4.6
A cached prefix on this model must be at least 4,096 tokens. Below that, Anthropic ignores cache_control silently — no error is returned and you pay full input price on
every request. A cache hit costs $0.5 per million tokens, roughly 10% of the
base input rate.
Comparable models from other providers
| Model | Provider | Input | Output |
|---|---|---|---|
| GPT-5.6 Sol | OpenAI | $5 | $30 |
| GPT-5.5 | OpenAI | $5 | $30 |
| GPT-5.4 | OpenAI | $2.5 | $15 |
| GPT-5.6 Terra | OpenAI | $2 | $12 |
All figures verified 2026-08-18 against Anthropic's official pricing documentation . Fields the provider does not publish are shown as “—” rather than estimated.