Proprietary
Claude Fable 5 — pricing, context window and what it's actually good for
Anthropic's most capable widely released model, positioned for long-running agents. 1M-token context, 128k max output. Reliable knowledge cutoff January 2026 (training data cutoff also January 2026). Claude 4.7 and later use a newer tokenizer that produces roughly 30% more tokens for the same text than earlier Claude models. Per-token price is therefore not comparable across that boundary — the same document costs about 30% more to process. Source: Anthropic pricing docs, checked 2026-08-07 11:51 UTC.
- Context window
- 1,000,000 tokens
- Max output
- 128,000 tokens
- Released
- 9 Jun 2026
- Knowledge cutoff
- Jan 2026
- API identifier
- claude-fable-5
Pricing · per 1M tokens
Cached input is 10× cheaper on this model — it pays off when you send the same context again and again.
US data residency · 1.1× list
Pinning inference to the US multiplies every token category — input, output, cache writes and cache reads — by 1.1. Global routing is the default and bills at the rates above. Anthropic pricing docs add that the partner-operated platforms — Amazon Bedrock and Google Cloud — have independent regional pricing, so check theirs rather than assuming these figures carry over.
Checked against Anthropic pricing docs on 7 Aug 2026, 11:51 UTC
Not sure Claude Fable 5 fits your workload?
We'll size the right model for your use case and give you a real cost estimate. One free call, no slides.
Talk to us