ClaudeHowSupport Us

claude-fable-5-1 — pricing, context window and limits

Anthropic's most capable widely released model: Fable 5's successor at the same per-token price, with cache reads at a quarter of Fable 5's. Thinking is always on, forced tool_choice returns a 400, and thinking blocks are bound to the model and conversation that produced them. Requires 30-day data retention, so zero-data-retention organisations get a 400.

Price per million tokens

Input$10
Output$50
Cache write, 5-minute TTL$12.50
Cache write, 1-hour TTL$20
Cache read$0.25 (0.025x input)
Batch input / output$5 / $25

Limits and behaviour

StatusCurrent
Context window1M tokens
Max output128K tokens
Min cacheable prefix512 tokens
Effort levelslow, medium, high, xhigh, max
Effort defaulthigh
ThinkingAlways on — cannot be disabled
Forced tool_choiceRejected (400)
PlatformsClaude API, Amazon Bedrock, Claude Platform on AWS, Google Cloud, Microsoft Foundry
Released1 Sept 2026

Lifecycle

API model nameclaude-fable-5-1
Retires no sooner than1 Sept 2027

The top of the lineup, with the cheapest cache reads in it

Fable 5.1 replaced Fable 5 in the top tier at the same per-token price for input and output. The change that moves a real bill is the cache: a read costs a fraction of Fable 5's rate, the deepest cache discount of any Claude model. On a long agentic session that re-reads a large prefix every turn, most of the input bill is cache reads, so this one rate outweighs the headline price.

A deeper read discount has a side effect worth planning for: a cache miss now costs far more relative to a hit. Anthropic's own guidance for idle gaps that outlast the short TTL but not the long one is that a cheap keep-alive request on the default short TTL usually beats paying the pricier write for the longer TTL. See cache reads got cheaper on Fable 5.1 and Opus 5.5.

The breaking changes from Fable 5

Forced tool choice (any or tool) returns a 400. Thinking blocks are bound to the model that produced them — only Mythos 5.1 can read a Fable 5.1 block, so a fallback to another model runs without that reasoning. And editing an earlier turn invalidates every later thinking block; for accounts created on or after the enforcement date that edit is a 400. See thinking block bound to a different conversation.

Who it is for

Anthropic positions Opus 5.5 as the default and Fable 5.1 as the step up for the hardest long-running agentic and research work, or where Opus 5.5 at higher effort still falls short. It requires Anthropic's data-retention window, so zero-retention organisations cannot call it, and it has no Priority Tier.

See migrating: Fable 5 to Fable 5.1 and Opus 5.5 vs Fable 5.1.

Verified 2026-09-30 against claude-api skill — shared/models.md, model-migration.md + SKILL.md model table and Anthropic's pricing and model-deprecation pages (canonical: https://platform.claude.com/docs/en/about-claude/models/overview, https://platform.claude.com/docs/en/about-claude/pricing, https://platform.claude.com/docs/en/about-claude/model-deprecations).Could not confirm: Microsoft Foundry availability is stated for Anthropic-hosted deployments only, and the 1M context window on Amazon Bedrock and Google Cloud was still open at launch, so we do not promise either.

Verified 2026-09-30 against Anthropic's model-deprecations status table and pricing page (canonical: https://platform.claude.com/docs/en/about-claude/model-deprecations, https://platform.claude.com/docs/en/about-claude/pricing).Could not confirm: These dates apply to the Claude API, Claude Platform on AWS and Microsoft Foundry. Amazon Bedrock and Google Cloud keep their own retirement schedules, which we do not track.