Claude Opus 4.7 pricing

Claude Opus 4.7 costs $5 per million input tokens and $25 per million output tokens , with cached input at $0.5. Its context window is 1M tokens. Batch requests are discounted 50%.

Retirement scheduled. Anthropic lists this model for shutdown on 2027-04-16 (a "not sooner than" floor — at least 60 days of notice precedes the real date). The recommended replacement is claude-opus-5. See all shutdown dates →

Price per million tokens

Input $5
Cached input $0.5
Output $25
Context window 1M tokens
Batch discount −50%
Minimum prefix to cache 2,048 tokens
Tokenizer factor ×1.3

What it costs in practice

Monthly cost for three common workload shapes, with prompt caching applied.

Workload Shape Per request Per month
Support chatbot 1.5k in / 400 out, 100k requests/month, 30% cached $0.0201 $2,012
RAG / document Q&A 12k in / 700 out, 50k requests/month, 60% cached $0.0586 $2,931
Coding agent 25k in / 2.5k out, 20k requests/month, 80% cached $0.127 $2,535

Model your own workload with caching, batching and tokenizer correction →

Caching on Claude Opus 4.7

A cached prefix on this model must be at least 2,048 tokens. Below that, Anthropic ignores cache_control silently — no error is returned and you pay full input price on every request. A cache hit costs $0.5 per million tokens, roughly 10% of the base input rate.

Check whether your system prompt clears the threshold →

Comparable models from other providers

Model Provider Input Output
GPT-5.6 Sol OpenAI $5 $30
GPT-5.5 OpenAI $5 $30
GPT-5.4 OpenAI $2.5 $15
GPT-5.6 Terra OpenAI $2 $12

All figures verified 2026-08-18 against Anthropic's official pricing documentation . Fields the provider does not publish are shown as “—” rather than estimated.