What Is Claude Sonnet 5.5?

What is Claude Sonnet 5.5? Anthropic's Sept 28, 2026 model: $2/$10 pricing, 1M context, benchmarks, what changed from Sonnet 5, and where to use it.

Ashley Goolam

Ashley Goolam

29 September 2026

What Is Claude Sonnet 5.5?

Apidog for Enterprise

On-Premises Deploy

SSO & RBAC

SOC 2 Compliant

Explore Apidog Enterprise

Claude Sonnet 5.5 is Anthropic’s mid-tier model, released on September 28, 2026, with the API id claude-sonnet-5-5. It costs $2 per million input tokens and $10 per million output tokens, the same as Sonnet 5. In its launch post, Anthropic says it “runs 30%+ faster” than Sonnet 5 and, in Anthropic’s testing, “costs up to 30% less per task than its predecessor.” It’s the second model in the Claude 5.5 family: a cheaper complement to Opus 5.5 ($4/$20) and a step above Haiku 4.5.

Below: specs, the lineup, changes from Sonnet 5, benchmarks, where to run it, and who should switch, with links to deeper guides like the pricing breakdown. To call the model as you read, Apidog can send the Messages request, keep your key in an environment variable, and show the response and token usage together.

Claude Sonnet 5.5 specs at a glance

Spec Claude Sonnet 5.5
Release date September 28, 2026
Model id (Claude API, Google Cloud, Microsoft Foundry, Claude Platform on AWS) claude-sonnet-5-5
Model id (Amazon Bedrock) anthropic.claude-sonnet-5-5
Context window 1M tokens
Max output 128K; 300K on the Batch API with the output-300k-2026-03-24 beta
Thinking Adaptive, on by default
Default effort high on the Claude Platform; medium in Claude Code and the Claude apps
Knowledge cutoff June 2026
Retirement Not sooner than September 28, 2027
Price per MTok $2 input, $10 output, $0.20 cache read, $2.50 (5-minute) or $4 (1-hour) cache write; Batch $1 / $5

Sources: the Sonnet 5.5 model overview and Claude pricing page.

There’s no long-context premium: a 900K-token request costs the same per token as a 9K one. Fast mode isn’t available; it’s Opus-only. US-only inference (inference_geo: "us") costs 1.1x. Sonnet 5’s planned September 1 rise to $3/$15 was cancelled, so $2/$10 is the standard Sonnet rate.

Where Sonnet 5.5 sits in the Claude lineup

Model Input / output per MTok Context / max output On Claude Free? Role
Claude Fable 5.1 $10 / $50 1M / 128K No Demanding reasoning, long-horizon agentic work
Claude Opus 5.5 $4 / $20 1M / 128K No Complex, open-ended work
Claude Sonnet 5.5 $2 / $10 1M / 128K Yes (chat) Well-scoped everyday work
Claude Haiku 4.5 $1 / $5 200K / 64K Yes Lowest-cost tier

Anthropic pitches Sonnet 5.5 for well-scoped everyday tasks, bug fixes, and polished docs, slides and spreadsheets, and says Opus 5.5 “remains clearly stronger at complex, open-ended work.” Haiku 5.5 is due “in the coming weeks.” For the tier above, read what Claude Opus 5.5 is.

What’s new compared with Sonnet 5

Price and tokenizer are unchanged. The API contract isn’t. The migration guide lists five breaking changes:

  1. thinking: {"type": "disabled"} is gone. Use {"type": "between_tools"}, accepted only at low, medium or high effort.
  2. Forced tool use is removed. tool_choice of any or tool fails. Send auto with strict: true tools and say in the prompt when to use them.
  3. Thinking blocks are bound to the conversation. For accounts created on or after 2026-08-31, replaying a block after editing earlier history fails. Keep conversations append-only.
  4. computer_20251124 is rejected on the Claude API and Google Cloud. Use computer_toolset_20260801.
  5. Advisor pairings are restricted to Opus 5, Opus 5.5, Sonnet 5.5, Fable 5, Fable 5.1, Mythos 5 or Mythos 5.1.

A sixth change fails nothing but alters the response shape: text between tool calls now arrives as progress-update thinking blocks, empty under the default display: "omitted", so your UI can go quiet mid-task. Set display: "updates" (beta header thinking-display-updates-2026-08-18) or "summarized" to see them.

Also new:

The Sonnet 5.5 vs Sonnet 5 guide has the exact errors and before/after JSON for each fix.

Claude Sonnet 5.5 benchmarks

The launch table compares Sonnet 5.5 with Sonnet 5, Opus 5.5 and GPT-6 Sol. Sonnet 5.5 ran at max effort unless a row says otherwise.

Benchmark Sonnet 5.5 Sonnet 5 Opus 5.5 GPT-6 Sol
Terminal-Bench 4.0 70.6% 10.3% 66.4% (xhigh) n/r
FrontierCode 1.1 Main (max) 46.2% 42.4% 54.4% 49.3%
FrontierCode 1.1 Main (xhigh) 52.1% 42.7% 51.4% 48.4%
CursorBench 4.0 55.5% 34.1% 57.8% n/r
GDPval-AA v2.1 (Elo) 1844 1449 1846 1487
Humanity’s Last Exam (tools) 64.5% 54.9% 67.7% n/r
OSWorld 2.1 (partial) 80.1% 57.0% 81.8% n/r

n/r: not reported. The xhigh FrontierCode row is from the launch page’s per-effort chart.

Anthropic ran Terminal-Bench, Humanity’s Last Exam and OSWorld itself. Cognition ran FrontierCode, where Sonnet 5.5 scored lower at max: it more often fanned out review subagents, which in cases Cognition examined caused a timeout or out-of-scope edits. Cursor ran CursorBench; Artificial Analysis ran GDPval-AA. Read the Terminal-Bench lead over Opus with care: Artificial Analysis measured Sonnet 5.5 at 63.6% in its own run, and the system card says fallback touched 10% of Opus trials versus 1.5% of Sonnet’s.

Sonnet 5.5 scores 56 on the Artificial Analysis Intelligence Index, #3 of 216, behind two Opus 5.5 endpoints (Sonnet 5 scored 38). The catch is verbosity: the index cost $8,977 to run at max effort versus $1,176 at high. The benchmarks deep dive has the per-effort tables.

Where you can use Claude Sonnet 5.5

Surface Access Notes
Claude API claude-sonnet-5-5 Prepaid credits
Amazon Bedrock anthropic.claude-sonnet-5-5 Global cross-Region inference only (commercial Regions); no strict tools
Google Cloud, Microsoft Foundry, Claude Platform on AWS claude-sonnet-5-5 Same id as the Claude API
Claude.ai (web, iOS, Android) Chat Anyone, including Free
Claude Code claude --model claude-sonnet-5-5 Paid plans, v2.1.284+; default stays Opus 5.5
GitHub Copilot VS Code, JetBrains, Copilot CLI and more Pro, Pro+, Max, Business, Enterprise
Cursor claude-sonnet-5-5 Other Models pool (Pro and above)
OpenRouter, Vercel AI Gateway anthropic/claude-sonnet-5.5 $2 / $10; no free variant

In Claude Code, the sonnet alias maps to Sonnet 5.5 only on the Anthropic API; on Bedrock, Google Cloud and Foundry it still resolves to Sonnet 4.5, so pin the full id. See Sonnet 5.5 in Claude Code for plans and effort. GitHub’s Copilot changelog says the rollout is gradual.

Anyone can chat with Sonnet 5.5 on the Claude Free plan, but a chat plan includes no API key, and API usage is prepaid. See how to use Sonnet 5.5 for free and whether a free Sonnet 5.5 API exists.

Who should switch to Sonnet 5.5

What early customers report

Customer claims from Anthropic’s launch page, not independent tests:

Send your first Sonnet 5.5 request

curl https://api.anthropic.com/v1/messages \
  -H "x-api-key: $ANTHROPIC_API_KEY" \
  -H "anthropic-version: 2023-06-01" \
  -H "content-type: application/json" \
  -d '{
    "model": "claude-sonnet-5-5",
    "max_tokens": 16000,
    "output_config": {"effort": "medium"},
    "messages": [
      {"role": "user", "content": "List three risks of retrying a POST request without an idempotency key."}
    ]
  }'

Thinking is on by default, so skip the thinking field. Leave out temperature, top_p and top_k: non-default values return a 400. Refusals come back as HTTP 200 with stop_reason: "refusal" and a stop_details category, so check the body, not only the status.

In Apidog, paste the curl command into a new request, move the key into an ANTHROPIC_API_KEY environment variable, and assert that stop_reason isn’t refusal or max_tokens. Duplicate it at low and high effort to compare usage. The Sonnet 5.5 API guide covers streaming, tools and caching; our Anthropic API key guide helps if you need a key.

FAQ

When was Claude Sonnet 5.5 released? September 28, 2026, six days after Opus 5.5.

What is the Claude Sonnet 5.5 context window? 1M tokens, with 128K max output (300K on the Batch API with a beta header). Long prompts cost the same per token as short ones.

Is Claude Sonnet 5.5 free? In the Claude.ai chat app, yes, including on the Free plan. The API is prepaid and Claude Code needs a paid plan. See every free route.

What is the Claude Sonnet 5.5 knowledge cutoff? June 2026.

Is Sonnet 5.5 the default model in Claude Code? No. The default stays Opus 5.5 on every paid plan. Switch with /model or claude --model claude-sonnet-5-5.

Next step

Take one workload you run on Sonnet 5 or Opus 5.5 today. Send it to claude-sonnet-5-5 at low, medium and high, compare quality and usage, and keep the cheapest setting that passes your checks. Download Apidog to keep those requests and assertions in one project, then use the pricing breakdown to turn token counts into a monthly bill.

Explore more

Generate API Tests With GPT-6 Luna: What a Full OpenAPI Spec Actually Costs

Generate API Tests With GPT-6 Luna: What a Full OpenAPI Spec Actually Costs

A full test-generation pass over a 42-endpoint OpenAPI spec costs about $0.29 on GPT-6 Luna at $0.10/$0.50, or under $0.10 with prompt caching. The per-spec arithmetic, the request shape, the latency to budget for, and the two failure modes.

23 September 2026

What Is GPT-6 Sol? Model ID, Pricing, 872K Context, and Benchmarks

What Is GPT-6 Sol? Model ID, Pricing, 872K Context, and Benchmarks

GPT-6 Sol explained: model ID gpt-6-sol, $2/$10 pricing, 872K context, AutomationBench 33.2% at $0.27 per task, DeepSWE 68.8%, and why it is a new model, not the GPT-5.6 Sol tier.

23 September 2026

What Is Claude Opus 5.5?

What Is Claude Opus 5.5?

Claude Opus 5.5 explained: model id claude-opus-5-5, $4/$20 pricing, 1M context, 128k max output, the eight benchmark scores Anthropic published, and 18+ hour tasks.

23 September 2026

Practice API Design-first in Apidog

Discover an easier way to build and use APIs

What Is Claude Sonnet 5.5?