Proprietary
Claude Opus 4.7 — pricing, context window and what it's actually good for
Claude Opus 4.7 is an Active Claude model on the Claude API. 1M-token context, 128k max output. Training data cutoff January 2026. Uses the tokenizer introduced with Claude Opus 4.7, which produces roughly 30% more tokens for the same text than Claude Sonnet 4.6 and earlier. Per-token price is not comparable across that boundary. Tentative retirement date: not sooner than 16 April 2027 (Anthropic model deprecations, checked 2026-08-07 11:51 UTC).
- Context window
- 1,000,000 tokens
- Max output
- 128,000 tokens
- Released
- 16 Apr 2026
- Knowledge cutoff
- Jan 2026
- API identifier
- claude-opus-4-7
Pricing · per 1M tokens
Cached input is 10× cheaper on this model — it pays off when you send the same context again and again.
US data residency · 1.1× list
Pinning inference to the US multiplies every token category — input, output, cache writes and cache reads — by 1.1. Global routing is the default and bills at the rates above. Anthropic pricing docs add that the partner-operated platforms — Amazon Bedrock and Google Cloud — have independent regional pricing, so check theirs rather than assuming these figures carry over.
Checked against Anthropic pricing docs on 7 Aug 2026, 11:51 UTC
Not sure Claude Opus 4.7 fits your workload?
We'll size the right model for your use case and give you a real cost estimate. One free call, no slides.
Talk to us