Proprietary
Claude Sonnet 5 — pricing, context window and what it's actually good for
Anthropic's balance of speed and intelligence. 1M-token context, 128k max output. Reliable knowledge cutoff January 2026. Introductory pricing of $2/$10 per 1M tokens runs through 31 August 2026, after which standard $3/$15 applies. Claude 4.7 and later use a newer tokenizer that produces roughly 30% more tokens for the same text than earlier Claude models. Per-token price is therefore not comparable across that boundary — the same document costs about 30% more to process. Source: Anthropic pricing docs, checked 2026-08-07 11:51 UTC.
- Context window
- 1,000,000 tokens
- Max output
- 128,000 tokens
- Released
- 30 Jun 2026
- Knowledge cutoff
- Jan 2026
- API identifier
- claude-sonnet-5
Pricing · per 1M tokens
Cached input is 10× cheaper on this model — it pays off when you send the same context again and again.
US data residency · 1.1× list
Pinning inference to the US multiplies every token category — input, output, cache writes and cache reads — by 1.1. Global routing is the default and bills at the rates above. Anthropic pricing docs add that the partner-operated platforms — Amazon Bedrock and Google Cloud — have independent regional pricing, so check theirs rather than assuming these figures carry over.
Checked against Anthropic pricing docs on 7 Aug 2026, 11:51 UTC
Introductory pricing of $2/$10 per 1M input/output tokens applies through 2026-08-31; standard $3/$15 takes effect 2026-09-01. The figures stored here are the STANDARD rates. Any cost example must state which rate it used.
Used for
Where Claude Sonnet 5 shows up in practice
The workflows where we shortlist Claude Sonnet 5 against the alternatives. Each page says why, at what price, and where a cheaper model does the job instead.
Not sure Claude Sonnet 5 fits your workload?
We'll size the right model for your use case and give you a real cost estimate. One free call, no slides.
Talk to us