Claude Sonnet 5.5 Pricing: The Full Cost Breakdown (API, Caching, Batch, and Plans)

Claude Sonnet 5.5 pricing: $2/$10 per million tokens, $0.20 cache reads, batch at half price. Worked cost examples, effort costs, and plan prices.

Ashley Goolam

Ashley Goolam

29 September 2026

Claude Sonnet 5.5 Pricing: The Full Cost Breakdown (API, Caching, Batch, and Plans)

Apidog for Enterprise

On-Premises Deploy

SSO & RBAC

SOC 2 Compliant

Explore Apidog Enterprise

Claude Sonnet 5.5 costs $2 per million input tokens and $10 per million output tokens on the Claude API, the same as Sonnet 5. Cache reads cost $0.20 per million, the Batch API halves both rates to $1 and $5, the full 1M context window bills at the standard rate, and fast mode isn’t offered. Sonnet 5’s planned September 1 rise to $3/$15 was cancelled, so $2/$10 is now the standard price.

This guide lists every line item of Claude Sonnet 5.5 pricing, works through three cost examples with the arithmetic shown, and explains why effort moves your bill more than the per-token rate. New to the model? Start with what Claude Sonnet 5.5 is, or read how to use Claude Sonnet 5.5 for free. To check the numbers against your own traffic, send the request from Apidog and read the usage block on each response.

Claude Sonnet 5.5 API pricing table

Prices per million tokens (MTok), from Anthropic’s pricing docs:

Line item Price per MTok
Input $2.00
5-minute cache write $2.50
1-hour cache write $4.00
Cache read $0.20
Output $10.00
Batch input $1.00
Batch output $5.00

Three details change what you pay:

How Sonnet 5.5 compares on list price

Model Input Output Notes
Claude Sonnet 5.5 $2 $10 Cache read $0.20; 1M context; no fast mode
Claude Sonnet 5 $2 $10 Cache read $0.20; $3/$15 rise cancelled
Claude Opus 5.5 $4 $20 Cache read $0.20; fast mode $8/$40
Claude Haiku 4.5 $1 $5 200k context
GPT-6 Sol $2 $10 90% off cached input; 872k context
GPT-6 Luna $0.10 $0.50 1M context

Sonnet 5.5 costs half of Opus 5.5 on input, output and cache writes. It shares the exact $2/$10 list price with GPT-6 Sol, so choosing between those two comes down to results per task. For the backstory, see Claude Sonnet 5 pricing, the September 2026 price war and GPT-6 Sol pricing.

Worked examples: what real requests cost

The formula for every line: tokens ÷ 1,000,000 × price per MTok. Price each line item, then add.

Example 1: one chat-style request

3,000 input tokens, 800 output tokens, no caching.

A thousand of those cost $14. Output is 21% of the tokens but 57% of the bill, so output length matters more than prompt length.

Example 2: an agent re-reading a 50,000-token prefix

An agent loop sends the same 50,000-token prefix (system prompt, tools, repo context) on 20 calls. Only the prefix is priced here.

Uncached: 20 × 50,000 = 1,000,000 tokens; 1,000,000 ÷ 1,000,000 × $2 = $2.00

Cached, with a 5-minute write:

That saves $1.685, about 84%. The 5-minute write pays off only while calls land inside the window; for slower loops, a 1-hour write costs 50,000 ÷ 1,000,000 × $4 = $0.20, so the total is $0.20 + $0.19 = $0.39, still about 80% below uncached. The same reasoning is worked for Opus in our prompt caching cost math; the $0.20 read rate is identical.

Example 3: a batch job of 10,000 requests

Same shape as Example 1, sent through the Batch API.

Batch also raises the output ceiling from 128K to 300K tokens with the output-300k-2026-03-24 beta.

Cost per task, not per token

The rate card tells you what a token costs, not how many tokens your task burns. That second number swings far more with effort than any rate gap between models.

Anthropic says that in its testing Sonnet 5.5 “costs up to 30% less per task than its predecessor.” The launch post also plots cost at every effort level. The Terminal-Bench 4.0 runs are Anthropic’s own, FrontierCode was run by Cognition, and the index costs come from Artificial Analysis:

Effort Terminal-Bench 4.0 (score, $/attempt) FrontierCode v1.1 (score, $/task) AA index (score, cost to run)
low 20.0%, $0.76 29.3%, $0.19 35.8, $544
medium 28.8%, $0.83 36.5%, $0.24 40.7, $701
high 43.0%, $1.94 49.4%, $0.42 46.7, $1,176
xhigh 61.5%, $5.30 52.1%, $1.59 51.9, $2,738
max 70.6%, $12.54 46.2%, $20.78 56.0, $8,977

What stands out:

Verbosity at max is the other trap. OfficeChai’s write-up of Artificial Analysis data puts Sonnet 5.5 at about 193,000 output tokens per index task at max, with cost per task about 50% above Sonnet 5. Simon Willison reported on Hacker News that a max-effort run burned 128,000 thinking tokens and ran out before producing the answer.

Six levers that cut the bill

  1. Set effort on purpose. The API defaults to high; Claude Code and the Claude apps default to medium. Anthropic’s prompting guide suggests medium or low for chat and xhigh or max only for measured quality gains. Don’t carry Sonnet 5 settings over; the scale was recalibrated.
  2. Use between_tools where you had thinking off. thinking: {"type": "disabled"} now returns a 400; {"type": "between_tools"} replaces it at low, medium and high.
  3. Cache anything reused. The minimum cacheable prompt is 512 tokens (1,024 on Sonnet 5). Changing top-level effort between requests invalidates the cache, so hold it steady.
  4. Batch what can wait. Half price on input and output.
  5. Keep max_tokens honest. Anthropic’s guide treats stop_reason: "max_tokens" as a failed response, so a truncated answer is spend with nothing to ship.
  6. Trim unrequested work. Sonnet 5.5 adds tests, docs and files you didn’t ask for, and starts extra review rounds at xhigh and max. Anthropic’s suggested system-prompt paragraph cut session cost by about a third at max with no quality change.

Where else you pay for Sonnet 5.5

Hunting for credits instead? See is there a free Claude Sonnet 5.5 API.

Track cost per request in Apidog

Every Messages response carries a usage object with input_tokens, output_tokens and cache counts. That’s enough to price any call in Apidog:

  1. Store your key as an ANTHROPIC_API_KEY environment variable.
  2. Save this request once per effort level (low, medium, high) with the same prompt:
curl https://api.anthropic.com/v1/messages \
  -H "x-api-key: $ANTHROPIC_API_KEY" \
  -H "anthropic-version: 2023-06-01" \
  -H "content-type: application/json" \
  -d '{
    "model": "claude-sonnet-5-5",
    "max_tokens": 4000,
    "thinking": {"type": "between_tools"},
    "output_config": {"effort": "medium"},
    "messages": [{"role": "user", "content": "Classify this support ticket: ..."}]
  }'
  1. Add a post-processor that prices the call and fails on truncation:
const body = pm.response.json();
const u = body.usage;
const cost = (u.input_tokens * 2 + u.output_tokens * 10) / 1e6;
pm.environment.set("last_call_usd", cost.toFixed(5));
pm.test("not truncated", () => pm.expect(body.stop_reason).to.not.eql("max_tokens"));
  1. Run all three and compare usage.output_tokens and cost side by side. If you cache, add the cache counts at their own rates.

The full request walkthrough is in how to use the Claude Sonnet 5.5 API.

FAQ

How much does Claude Sonnet 5.5 cost per million tokens? $2 input and $10 output on the Claude API. Cache reads cost $0.20; the Batch API charges $1 and $5.

Is Sonnet 5.5 more expensive than Sonnet 5? No. The per-token price is identical, and Anthropic says Sonnet 5.5 costs up to 30% less per task in its testing.

Is there a surcharge above 200k tokens? No. The full 1M window bills at the standard rate.

Does Sonnet 5.5 have fast mode? No. Fast mode is available only on Opus 5.5, Opus 5 and Opus 4.8.

Is Claude Sonnet 5.5 free? In the Claude chat app, yes: anyone can chat with it on the Free plan. The API runs on prepaid credits; the free API guide covers the credit programs.

Next step: price your own workload

Take your three most common request shapes, run each at two effort levels, and multiply the usage numbers by the table above. You’ll see quickly whether medium holds your quality bar or high earns its extra tokens. Download Apidog to keep those requests and the cost script side by side.

Explore more

How to Use Claude Sonnet 5.5 for Free: Every Route That Works (and the Ones That Don't)

How to Use Claude Sonnet 5.5 for Free: Every Route That Works (and the Ones That Don't)

Is Claude Sonnet 5.5 free? Yes on Claude.ai (web, iOS, Android). Every free route checked, plus what isn't: Claude Code, the API, and Copilot Free.

29 September 2026

How to Use Claude Sonnet 5.5 in Claude Code (and When to Keep Opus 5.5)

How to Use Claude Sonnet 5.5 in Claude Code (and When to Keep Opus 5.5)

Claude Sonnet 5.5 Claude Code setup: v2.1.284+, claude --model claude-sonnet-5-5, effort levels, the sonnet alias trap, and when to keep Opus 5.5.

29 September 2026

Claude Sonnet 5.5 vs Sonnet 5: What Changed, and the Breaking Changes to Fix Before You Switch

Claude Sonnet 5.5 vs Sonnet 5: What Changed, and the Breaking Changes to Fix Before You Switch

Sonnet 5.5 vs Sonnet 5: same $2/$10 price, far higher scores, and five breaking changes that return 400s. Exact errors and before/after JSON fixes.

29 September 2026

Practice API Design-first in Apidog

Discover an easier way to build and use APIs

Claude Sonnet 5.5 Pricing: The Full Cost Breakdown (API, Caching, Batch, and Plans)