█

AI/TLDR

Claude Sonnet 5.5

Anthropic's Sonnet-tier model of the Claude 5.5 family: over 30% faster than Sonnet 5 at the same $2/$10 per MTok, with 70.6% on Terminal-Bench 4.0.

Claude SonnetAPI onlyGenerally available
Released
28 Sep 2026
Context
1M
Input
$2.00 per million tokens
License
Proprietary
Coverage
1 story

Overview

Claude Sonnet 5.5 is the Sonnet-tier model of Anthropic's Claude 5.5 family, released on September 28, 2026, six days after Claude Opus 5.5. Anthropic describes it as the best combination of speed and intelligence in its lineup, and says it generates output more than 30% faster than Claude Sonnet 5 and costs up to 30% less per task.

Claude Sonnet 5.5's starling murmuration program running, a flock of arrow-shaped birds after 4,158 output tokens
From Anthropic's side-by-side demo: the newer model running the murmuration program it wrote…
The previous Claude model writing the boids loop of the same murmuration program, 4,520 tokens in
…and the previous model, shown beside it in the same demo, writing the same program.

The model has a 1 million-token context window and up to 128K output tokens on the synchronous Messages API, rising to 300K on the Message Batches API with the output-300k-2026-03-24 beta header. It takes text and images in and returns text. Reliable knowledge runs through June 2026. Adaptive thinking is on by default and its depth is set with the effort parameter, which defaults to high; the tokenizer is the same as Claude Sonnet 5's.

Pricing is unchanged from Claude Sonnet 5: $2 per million input tokens and $10 per million output tokens, half the $4/$20 of Claude Opus 5.5. Cache reads cost $0.20 per million, a 5-minute cache write $2.50 and a 1-hour cache write $4; Batch API requests take 50% off both sides. The minimum cacheable prompt is 512 tokens, down from 1,024 on Sonnet 5.

Claude Sonnet 5.5 is the first Sonnet model to launch with the cyber safeguards and fallbacks Anthropic built for its most capable models. Five changes break code written for Claude Sonnet 5: thinking type disabled returns an error (between_tools replaces it), forced tool use is rejected, thinking blocks are tied to the model and conversation that produced them, the computer_20251124 tool is refused on the Claude API and Google Cloud, and Opus 4.8, Opus 4.7 and Sonnet 5 no longer work as advisors.

Released2026-09-28
LicenseProprietary
WeightsAPI only
Context1M
Max output128K
ArchitectureProprietary transformer with adaptive thinking on by default, steered by an effort parameter; the lowest setting, between_tools, turns off up-front thinking.
Knowledge cutoffJun 2026
ModalitiesText, Vision
StatusGenerally available

Benchmarks

Claude Sonnet 5.5 vs Claude Sonnet 5 and Claude Opus 5.5

BenchmarkClaude Sonnet 5.5Claude Sonnet 5Claude Opus 5.5
Terminal-Bench 4.070.6%10.3%66.4%
FrontierCode 1.1 (Main)46.2%42.4%54.4%
CursorBench 4.055.5%34.1%57.8%
OSWorld 2.1 (partial)80.1%57%81.8%
Humanity's Last Exam (with tools)64.5%54.9%67.7%
GDPval-AA v2.1 (Elo)184414491846
AA-Briefcase v1.1 (Elo)181113591822

Comparison source ↗

This model's scores

  1. Terminal-Bench 4.0 (agentic coding)70.6%
  2. OSWorld 2.1 (partial)80.1%
  3. Humanity's Last Exam (with tools)64.5%
  4. CursorBench 4.055.5%
  5. FrontierCode 1.1 (Main, Max)46.2%

Scores on a 0–100 scale (25-point gridlines); higher is better. Each benchmark links to its published source.

Pricing

Input$2.00 per million tokens
Output$10.00 per million tokens

Cache reads $0.20/MTok, 5m cache write $2.50/MTok, 1h cache write $4.00/MTok. Batch API takes 50% off input and output.

Pricing source ↗

Strengths

  • 70.6% on Terminal-Bench 4.0 in Anthropic's launch comparison, ahead of Claude Opus 5.5 at 66.4% and Claude Sonnet 5 at 10.3%
  • Near-Opus computer use: 80.1% partial on OSWorld 2.1 against 81.8% for Opus 5.5 and 57.0% for Sonnet 5
  • Same $2/$10 per MTok as Sonnet 5, with output over 30% faster and up to 30% lower cost per task
  • 1M-token context with 128K max output, extendable to 300K on the Batch API
  • Supports per-message effort (beta), mid-conversation system messages and mid-conversation tool changes, which Sonnet 5 does not

Best for

  • Reach for it to run coding and terminal agents at volume, where near-flagship results at half the Opus 5.5 token price matter more than the last few points.
  • Reach for it for computer-use automation through the computer_toolset_20260801 toolset, where it scores within two points of Opus 5.5 on OSWorld 2.1.
  • Reach for it for latency-sensitive chat and iterative work, starting at medium or low effort as Anthropic recommends.

How to access

ProviderModel ID
Anthropic API ↗claude-sonnet-5-5
Amazon Bedrock ↗anthropic.claude-sonnet-5-5
Google Vertex AI ↗claude-sonnet-5-5
Microsoft Foundry ↗claude-sonnet-5-5

Claude Sonnet — every version

The full lineage of the Claude Sonnet line, newest first. Every version has its own page — click any to compare specs, benchmarks and pricing.

VersionReleasedContextLicense
Claude Sonnet 5.5current2026-09-281MProprietary
Claude Sonnet 52026-06-301MProprietary
Claude Sonnet 4.62026-02-171MProprietary
Claude Sonnet 4.52025-09-291MProprietary
Claude Sonnet 42025-05-22200KProprietary
Claude 3.7 Sonnet2025-02-24200KProprietary
Claude 3.5 Sonnet (new)2024-10-22200KProprietary
Claude 3.5 Sonnet2024-06-20200KProprietary
Claude 3 Sonnet2024-03-04200KProprietary

FAQ

When was Claude Sonnet 5.5 released?

Anthropic released Claude Sonnet 5.5 on September 28, 2026, six days after Claude Opus 5.5. It launched on the Claude API, Amazon Bedrock, Google Cloud, Microsoft Foundry and Claude Platform on AWS, and in claude.ai. Anthropic commits to keeping it available until at least September 28, 2027.

How much does Claude Sonnet 5.5 cost?

Claude Sonnet 5.5 costs $2 per million input tokens and $10 per million output tokens, the same as Claude Sonnet 5. Cache reads are $0.20 per million, a 5-minute cache write $2.50 and a 1-hour cache write $4. Batch API requests take 50% off. Anthropic says it still costs up to 30% less per task than Sonnet 5.

How does Claude Sonnet 5.5 compare to Claude Opus 5.5?

In Anthropic's launch comparison Claude Sonnet 5.5 beats Claude Opus 5.5 on Terminal-Bench 4.0, 70.6% to 66.4%, and trails it on OSWorld 2.1 (80.1% to 81.8%), CursorBench 4.0 (55.5% to 57.8%) and Humanity's Last Exam (64.5% to 67.7%). Sonnet 5.5 costs half as much per token: $2/$10 against $4/$20.

Can thinking be turned off on Claude Sonnet 5.5?

Claude Sonnet 5.5 rejects thinking type disabled with a 400 error. Its lowest setting is between_tools, which turns off up-front thinking while keeping short progress notes between tool calls. between_tools works at low, medium and high effort; at xhigh or max effort Claude Sonnet 5.5 requires adaptive thinking instead.

What breaks when migrating from Claude Sonnet 5 to Claude Sonnet 5.5?

Five changes break Claude Sonnet 5 code: thinking type disabled returns an error, forced tool use is rejected, thinking blocks are tied to the model and conversation, the computer_20251124 tool is refused on the Claude API and Google Cloud, and Opus 4.8, Opus 4.7 and Sonnet 5 advisors fail. Anthropic's migration guide covers each change.