Claude Opus 4.5 pricing

Claude Opus 4.5 costs $5 per million input tokens and $25 per million output tokens , with cached input at $0.5. Its context window is 200K tokens. Batch requests are discounted 50%.

Retirement scheduled. Anthropic lists this model for shutdown on 2026-11-24 (a "not sooner than" floor — at least 60 days of notice precedes the real date). The recommended replacement is claude-opus-5. See all shutdown dates →

Price per million tokens

Input $5
Cached input $0.5
Output $25
Context window 200K tokens
Batch discount −50%
Minimum prefix to cache 4,096 tokens

What it costs in practice

Monthly cost for three common workload shapes, with prompt caching applied.

Workload Shape Per request Per month
Support chatbot 1.5k in / 400 out, 100k requests/month, 30% cached $0.0155 $1,548
RAG / document Q&A 12k in / 700 out, 50k requests/month, 60% cached $0.0451 $2,255
Coding agent 25k in / 2.5k out, 20k requests/month, 80% cached $0.0975 $1,950

Model your own workload with caching, batching and tokenizer correction →

Caching on Claude Opus 4.5

A cached prefix on this model must be at least 4,096 tokens. Below that, Anthropic ignores cache_control silently — no error is returned and you pay full input price on every request. A cache hit costs $0.5 per million tokens, roughly 10% of the base input rate.

Check whether your system prompt clears the threshold →

Comparable models from other providers

Model Provider Input Output
GPT-5.6 Sol OpenAI $5 $30
GPT-5.5 OpenAI $5 $30
GPT-5.4 OpenAI $2.5 $15
GPT-5.6 Terra OpenAI $2 $12

All figures verified 2026-08-18 against Anthropic's official pricing documentation . Fields the provider does not publish are shown as “—” rather than estimated.