claude-fable-5-1 — pricing, context window and limits
Anthropic's most capable widely released model: Fable 5's successor at the same per-token price, with cache reads at a quarter of Fable 5's. Thinking is always on, forced tool_choice returns a 400, and thinking blocks are bound to the model and conversation that produced them. Requires 30-day data retention, so zero-data-retention organisations get a 400.
Price per million tokens
| Input | $10 |
|---|---|
| Output | $50 |
| Cache write, 5-minute TTL | $12.50 |
| Cache write, 1-hour TTL | $20 |
| Cache read | $0.25 (0.025x input) |
| Batch input / output | $5 / $25 |
Limits and behaviour
| Status | Current |
|---|---|
| Context window | 1M tokens |
| Max output | 128K tokens |
| Min cacheable prefix | 512 tokens |
| Effort levels | low, medium, high, xhigh, max |
| Effort default | high |
| Thinking | Always on — cannot be disabled |
| Forced tool_choice | Rejected (400) |
| Platforms | Claude API, Amazon Bedrock, Claude Platform on AWS, Google Cloud, Microsoft Foundry |
| Released | 1 Sept 2026 |
Lifecycle
| API model name | claude-fable-5-1 |
|---|---|
| Retires no sooner than | 1 Sept 2027 |
The top of the lineup, with the cheapest cache reads in it
Fable 5.1 replaced Fable 5 in the top tier at the same per-token price for input and output. The change that moves a real bill is the cache: a read costs a fraction of Fable 5's rate, the deepest cache discount of any Claude model. On a long agentic session that re-reads a large prefix every turn, most of the input bill is cache reads, so this one rate outweighs the headline price.
A deeper read discount has a side effect worth planning for: a cache miss now costs far more relative to a hit. Anthropic's own guidance for idle gaps that outlast the short TTL but not the long one is that a cheap keep-alive request on the default short TTL usually beats paying the pricier write for the longer TTL. See cache reads got cheaper on Fable 5.1 and Opus 5.5.
The breaking changes from Fable 5
Forced tool choice (any or tool) returns a 400. Thinking blocks are bound to the model that
produced them — only Mythos 5.1 can read a Fable 5.1 block, so a fallback to another model runs
without that reasoning. And editing an earlier turn invalidates every later thinking block; for
accounts created on or after the enforcement date that edit is a 400. See
thinking block bound to a different conversation.
Who it is for
Anthropic positions Opus 5.5 as the default and Fable 5.1 as the step up for the hardest long-running agentic and research work, or where Opus 5.5 at higher effort still falls short. It requires Anthropic's data-retention window, so zero-retention organisations cannot call it, and it has no Priority Tier.
Related
See migrating: Fable 5 to Fable 5.1 and Opus 5.5 vs Fable 5.1.
Verified 2026-09-30 against claude-api skill — shared/models.md, model-migration.md + SKILL.md model table and Anthropic's pricing and model-deprecation pages (canonical: https://platform.claude.com/docs/en/about-claude/models/overview, https://platform.claude.com/docs/en/about-claude/pricing, https://platform.claude.com/docs/en/about-claude/model-deprecations).Could not confirm: Microsoft Foundry availability is stated for Anthropic-hosted deployments only, and the 1M context window on Amazon Bedrock and Google Cloud was still open at launch, so we do not promise either.
Verified 2026-09-30 against Anthropic's model-deprecations status table and pricing page (canonical: https://platform.claude.com/docs/en/about-claude/model-deprecations, https://platform.claude.com/docs/en/about-claude/pricing).Could not confirm: These dates apply to the Claude API, Claude Platform on AWS and Microsoft Foundry. Amazon Bedrock and Google Cloud keep their own retirement schedules, which we do not track.