ClaudeHowSupport Us

Opus 5.5 vs Sonnet 5.5

Side by side

Opus 5.5Sonnet 5.5
Input / MTok$4$2
Output / MTok$20$10
Cache read / MTok$0.20$0.20
Context window1M1M
Max output128K128K
Min cacheable prefix512512
Effort levelslow, medium, high, xhigh, maxlow, medium, high, xhigh, max
Effort defaultmediumhigh
ThinkingAlways on — cannot be disabledOn by default — lowest setting is between_tools
Forced tool_choice400400

The pairing most teams now weigh

These are the two current models most production work chooses between. Anthropic positions Opus 5.5 as the default for most work, including complex agentic coding, and Sonnet 5.5 as the faster option for everyday coding, agent and enterprise work — with an Opus model the better choice for the hardest long-horizon tasks. Opus 5.5 costs more per token for input and output, as the table shows.

The detail that narrows the gap: cache reads

Read the cache row before the headline rows. Opus 5.5 discounts its cache reads more deeply than Sonnet 5.5 does, by enough that the two cache-read rates in the table come out the same. For an agentic workload where most input tokens are re-read from cache, the input side of the price gap largely disappears, and what remains is output plus fresh input. The effective difference on that kind of workload is much smaller than the headline suggests — run your own token mix through the token & cost estimator rather than trusting either impression.

Thinking controls differ

Opus 5.5 cannot switch thinking off at all; effort is the only dial, and it defaults to medium. Sonnet 5.5 defaults to high, and while it also rejects disabled, it offers between_tools as a thinking-free setting at effort high or below. For latency-critical routes that must not think, Sonnet 5.5 is the only one of the two with that option.

Choosing

For routine, well-scoped work, start on Sonnet 5.5 at a low or medium effort and move a route up only when an evaluation shows the gap. For long, multi-step work in a large codebase, start on Opus 5.5 at its default and test the neighbouring levels. Both reject forced tool choice, so the same request code serves either.

See the model picker and Sonnet 5.5 vs Haiku 4.5.

Verified 2026-09-30 against claude-api skill — shared/models.md, model-migration.md + SKILL.md model table and Anthropic's pricing and model-deprecation pages (canonical: https://platform.claude.com/docs/en/about-claude/models/overview, https://platform.claude.com/docs/en/about-claude/pricing, https://platform.claude.com/docs/en/about-claude/model-deprecations).Could not confirm: On Microsoft Foundry it is hosted on Azure only, so Foundry features that require Anthropic hosting (code execution, the Files API, the newer web tools) are unavailable for it there.