ClaudeHowSupport Us

Claude API pricing: every model, every rate

Standard rates, USD per million tokens

ModelInputOutputCache read
Fable 5.1$10$50$0.25 *
Mythos 5.1 †$10$50$0.25 *
Fable 5$10$50$1
Mythos 5 †$10$50$1
Opus 5.5$4$20$0.20 *
Opus 5$5$25$0.50
Opus 4.8$5$25$0.50
Opus 4.7$5$25$0.50
Opus 4.6$5$25$0.50
Sonnet 5.5$2$10$0.20
Sonnet 5$2$10$0.20
Sonnet 4.6$3$15$0.30
Haiku 4.5$1$5$0.10
Opus 4.5$5$25$0.50
Sonnet 4.5$3$15$0.30

* Cache reads below the standard 0.1x of input. † Invitation-only (Project Glasswing).

Cache writes and batch, USD per million tokens

ModelWrite, 5 minWrite, 1 hourBatch inBatch out
Fable 5.1$12.50$20$5$25
Mythos 5.1 †$12.50$20$5$25
Fable 5$12.50$20$5$25
Mythos 5 †$12.50$20$5$25
Opus 5.5$5$8$2$10
Opus 5$6.25$10$2.50$12.50
Opus 4.8$6.25$10$2.50$12.50
Opus 4.7$6.25$10$2.50$12.50
Opus 4.6$6.25$10$2.50$12.50
Sonnet 5.5$2.50$4$1$5
Sonnet 5$2.50$4$1$5
Sonnet 4.6$3.75$6$1.50$7.50
Haiku 4.5$1.25$2$0.50$2.50
Opus 4.5$6.25$10$2.50$12.50
Sonnet 4.5$3.75$6$1.50$7.50

Fast mode — Claude API only — not Claude Platform on AWS, Amazon Bedrock, Google Cloud or Microsoft Foundry

ModelInputOutputvs standard
Opus 5.5$8$402x
Opus 5$10$502x
Opus 4.8$10$502x

Modifiers and tool charges

Batch API0.5x standard on input and output; results within 24 hours
US-only inference (inference_geo)1.1x on Claude 4.6 and later models, on every token category including cache reads and writes
Fast mode speedUp to 2.5x output tokens per second; not with the Batch API or Priority Tier
Web search$10 per 1,000 searches, plus the tokens its results add
Web fetchNo charge beyond the tokens it adds
Code execution1,550 free hours a month per organisation, then $0.05 per container-hour; free alongside the newer web search and fetch tools
Managed Agents runtime$0.08 per session-hour while running, plus tokens

How to read these tables

Every rate above is per million tokens, in US dollars, as Anthropic publishes it for the first-party Claude API, and every one renders from this site's dated facts module rather than being typed onto the page. The same rates apply on the other Anthropic-operated platforms — Claude Platform on AWS and Microsoft Foundry — where the marketplace converts them into consumption units on your AWS or Azure invoice. Amazon Bedrock and Google Cloud set their own prices, which these tables do not cover.

The column that changed most recently is the cache read. Until Fable 5.1 and Opus 5.5, a cache read cost the same fraction of input on every model, so listing input and output was enough. It no longer is: the newest Fable and Opus models read their cache at a deeper discount, which is why a model's cache-read rate is printed beside its headline rate instead of being left for you to derive.

What the tables leave out, on purpose

A per-token rate is not a per-request cost. Three things sit between the two. The tokenizer: models from Opus 4.7 onward count more tokens for the same text than older ones, so identical prompts bill differently at identical rates. Thinking: reasoning tokens bill as output whether or not their text is returned to you, and on several current models thinking cannot be switched off. And tool use: a request that declares tools carries a small, model-specific system prompt, and server tools add their own charges, listed in the last table.

The newest models also bill their entire context window at the standard rate — there is no long-context surcharge on them — and caching and batch discounts stack across the full window.

Turning a rate into a bill

Paste real text into the token & cost estimator to price it on every model at once, use the prompt-caching savings calculator to see what the cache-read column is worth to your request pattern, and the batch vs realtime calculator for work that can wait. Each model's own reference page carries the same rates alongside its limits and lifecycle.

Verified 2026-09-30 against claude-api skill — shared/models.md, model-migration.md + SKILL.md model table and Anthropic's pricing and model-deprecation pages (canonical: https://platform.claude.com/docs/en/about-claude/models/overview, https://platform.claude.com/docs/en/about-claude/pricing, https://platform.claude.com/docs/en/about-claude/model-deprecations).

Verified 2026-09-30 against claude-api skill — shared/prompt-caching.md and Anthropic's pricing page (canonical: https://platform.claude.com/docs/en/build-with-claude/prompt-caching, https://platform.claude.com/docs/en/about-claude/pricing).

Verified 2026-09-30 against claude-api skill — {python,typescript}/claude-api/batches.md + shared/platform-availability.md and Anthropic's pricing page (canonical: https://platform.claude.com/docs/en/build-with-claude/batch-processing, https://platform.claude.com/docs/en/about-claude/pricing).

Verified 2026-09-30 against claude-api skill — SKILL.md Fast Mode quick reference and Anthropic's pricing page (canonical: https://platform.claude.com/docs/en/about-claude/pricing).

Verified 2026-09-30 against Anthropic's pricing page — data residency section (canonical: https://platform.claude.com/docs/en/about-claude/pricing).

Verified 2026-09-30 against Anthropic's pricing page — tool pricing and Managed Agents sections (canonical: https://platform.claude.com/docs/en/about-claude/pricing).