Skip links

Claude Sonnet 5.5 explained: same price, fewer tokens, one effort trap

Claude Sonnet 5.5, Anthropic's mid-tier model released on 28 September 2026, shown typed at a terminal prompt with agentic task rows running below
Sonnet 5.5 keeps Sonnet 5's $2 and $10 price, writes more than 30 per cent faster and finishes the same work in fewer tokens.

Anthropic released Claude Sonnet 5.5 on 28 September 2026 at the same price as the model it replaces: $2 per million input tokens and $10 per million output. Anthropic says it writes more than 30 per cent faster than Claude Sonnet 5, needs far fewer tokens for the same job, and costs up to 30 per cent less per task as a result.

For most businesses this is the Claude model you will use every day, and the trade is a good one. There is one catch, and it sits in the effort setting. Turned up to maximum, Sonnet 5.5 writes more output tokens per task than any model Artificial Analysis has measured, and it costs more per task than Claude Opus 5.5 while scoring lower. Leave it at medium or high and you get the upgrade at the old price; at high effort it sits within a point of GPT-6 Sol for the same cost per task.

Prices, model IDs and migration rules were checked against Anthropic’s launch page, Sonnet 5.5 overview, what’s new page, migration guide and pricing documentation on 29 September 2026. Independent figures come from Artificial Analysis, checked the same day.

The answer in brief

  • Claude Sonnet 5.5 is Anthropic’s mid-tier model, released on 28 September 2026 as the second model in the Claude 5.5 family, after Opus 5.5. Haiku 5.5 follows in the coming weeks.
  • It costs $2 per million input tokens and $10 per million output, unchanged from Sonnet 5. Cache reads are $0.20 and five-minute cache writes $2.50.
  • Anthropic says it generates output more than 30 per cent faster than Sonnet 5 and costs up to 30 per cent less per task, because it needs far fewer tokens for the same work. Early testers also reported fewer steps and tool calls.
  • On the Artificial Analysis Intelligence Index it scores 56 at maximum effort, second only to Opus 5.5 on 58, and 18 points above Sonnet 5.
  • The effort setting decides the bill. At high effort it beats Sonnet 5’s best score for about a fifth of the cost per task. At maximum it costs $7.60 per task, more than Opus 5.5 at $5.98.
  • Migrating an API integration from Sonnet 5 needs code changes: disabled thinking, forced tool use and the old computer-use tool all return errors, and edited conversation history fails on newer API accounts.
  • Try it on one everyday workflow at medium effort, compare tokens, steps and rework against Sonnet 5, then move traffic in stages. Keep Opus 5.5 for long, open-ended work.

What is Claude Sonnet 5.5, and where can you use it?

Anthropic positions Sonnet 5.5 as the faster, lower-cost complement to Opus 5.5. Opus 5.5 is built for complex work that needs sustained judgement. Sonnet 5.5 is for well-scoped everyday tasks: fixing bugs, drafting documents, building slides and spreadsheets, and iterating quickly on smaller jobs. Anthropic also says early testers noticed a sharper eye for design and clearer writing than the previous generation.

It takes text and images and returns text. Adaptive thinking is on by default and you steer its depth with the effort setting, which has five levels: low, medium, high, xhigh and max. The reliable knowledge cutoff is June 2026. Anthropic’s deprecation schedule says it will not retire the model sooner than 28 September 2027, and Sonnet 5 stays active until at least 30 June 2027, so nobody is being forced to move.

Claude Sonnet 5.5 at a glance

Claude Sonnet 5.5 key specifications (checked 29 September 2026)

If the table extends beyond the screen, scroll sideways to view all columns.

Spec Value Source
Best suited to Well-scoped everyday tasks, bug fixes, documents, slides and spreadsheets Anthropic launch page
API price $2 input / $10 output per million tokens Anthropic pricing
Cache read $0.20 per million tokens Anthropic pricing
5-minute cache write $2.50 per million tokens Anthropic pricing
1-hour cache write $4 per million tokens Anthropic pricing
Batch API 50 per cent off input and output Anthropic pricing
Context / max output 1M tokens / 128K tokens (300K on the Batch API beta) Anthropic docs
Model ID claude-sonnet-5-5 (Bedrock: anthropic.claude-sonnet-5-5) Anthropic docs
Default effort high on the Claude API; medium in the Claude apps and Claude Code Anthropic docs and launch page
Knowledge cutoff June 2026 Anthropic docs
Earliest retirement 28 September 2027 Anthropic deprecations

Where you can use it

Anthropic lists Sonnet 5.5 on the Claude Platform, Amazon Bedrock, Google Cloud, Microsoft Foundry and Claude Platform on AWS, with zero data retention available as on Opus 5.5 and Sonnet 5. Its launch page says the model is in Claude Code and the Claude apps, where the default effort is medium. Anthropic’s pages do not say which plans get it as the default model, so check your model picker rather than assuming.

As with Opus 5.5, a Claude subscription does not mean your own application uses the API model. Check the product you intend to use, and review zero-data-retention and processor terms under UK GDPR before client data goes through it. Anthropic bills in US dollars, so convert at your own rate and add VAT where it applies.

How much does Claude Sonnet 5.5 cost?

Per million tokens, from Anthropic’s pricing page (checked 29 September 2026). Cache write rates depend on the cache lifetime, so both are shown:

Claude model API pricing comparison (USD per million tokens)

If the table extends beyond the screen, scroll sideways to view all columns.

Model Input Output Cache read 5m cache write 1h cache write
Claude Fable 5.1 $10 $50 $0.25 $12.50 $20
Claude Opus 5.5 $4 $20 $0.20 $5 $8
Claude Sonnet 5.5 $2 $10 $0.20 $2.50 $4
Claude Sonnet 5 $2 $10 $0.20 $2.50 $4
Claude Haiku 4.5 $1 $5 $0.10 $1.25 $2
Price cards showing Claude Sonnet 5.5 at $2 input and $10 output per million tokens with cache reads at $0.20, unchanged from Claude Sonnet 5, and Claude Opus 5.5 at $4 and $20
Figure 1. Same rate card as Sonnet 5; the saving Anthropic claims comes from fewer tokens per task, not a lower price.

The list price has not moved. Sonnet 5 launched at $2 and $10 as an introductory rate, Anthropic made that permanent on 10 August 2026, and Sonnet 5.5 inherits it, prompt caching and batch rates included. GPT-6 Sol sits at the same $2 and $10 on OpenAI’s API, so the price comparison between the two mid-tier models comes down to tokens per task, not the rate card.

That is where Anthropic’s saving claim lives. Its own testing puts the cost per task at up to 30 per cent below Sonnet 5, because the model finishes the same work in fewer tokens, and early testers saw fewer tool calls too. The customer figures on the launch page point the same way: Slack reports about 14 per cent fewer output tokens on its Slackbot evaluations, Box 12 per cent fewer total tokens than its previous model, Lovable a third fewer tool calls, and Balyasny Asset Management about 121,000 tokens per answer on its finance tasks against 497,000 for Sonnet 5. Those are vendor-selected testimonials, and none of them is a marketing workload, so measure your own.

Two smaller notes for budgets. Cache writes at the one-hour rate are $4 per million tokens, not $2.50, so agents that keep a long-lived cache pay more up front. And on the Batch API, Sonnet 5.5 costs $1 input and $5 output, which is the Haiku 4.5 list price for a much stronger model, if the job can wait.

Which effort setting should you use with Claude Sonnet 5.5?

Medium for routine work and high for anything that needs checking. Avoid max unless a side-by-side test proves it earns its cost. The independent numbers from launch week make the case.

Artificial Analysis ran Sonnet 5.5 at every effort level on its Intelligence Index v4.3.2, a composite of ten evaluations, with Anthropic’s default fallback switched on. The figures below were read from its model pages on 29 September 2026. Cost per task is Artificial Analysis’s weighted average across the index, at list prices:

Artificial Analysis Intelligence Index v4.3.2 by effort setting (checked 29 September 2026)

If the table extends beyond the screen, scroll sideways to view all columns.

Model and effort Index score Cost per task
Claude Sonnet 5.5, low 36 $0.41
Claude Sonnet 5.5, medium 41 $0.59
Claude Sonnet 5.5, high 47 $1.08
Claude Sonnet 5.5, max 56 $7.60
Claude Sonnet 5, max 38 $5.09
Claude Opus 5.5, medium 51 $1.34
Claude Opus 5.5, max 58 $5.98
GPT-6 Sol, max 48 $1.05
GPT-6 Astra, max 53 $3.26
Bar chart of cost per task on the Artificial Analysis Intelligence Index v4.3.2: Claude Sonnet 5.5 at high effort $1.08 with a score of 47, at max $7.60 scoring 56, and Claude Opus 5.5 at max $5.98 scoring 58
Figure 2. High effort beats Sonnet 5’s best score for about a fifth of the cost per task; maximum effort costs more than Opus 5.5 at max.

Read it from the bottom up. Sonnet 5.5 at medium beats Sonnet 5’s best score for about 12 per cent of the cost per task. At high it scores nine points above Sonnet 5’s best for about a fifth of the cost. That is the upgrade, and it arrives without a price change.

Then look at the top. At max, Sonnet 5.5 reaches 56, two points behind Opus 5.5, but it gets there by writing about 193,000 output tokens per task, the most Artificial Analysis has measured, around 60 per cent more than Opus 5.5 or Sonnet 5 at their maximum settings. The result is a cost per task of $7.60, against $5.98 for Opus 5.5 at max and $5.09 for Sonnet 5 at max. If a job needs Sonnet 5.5 at max, Opus 5.5 at max is cheaper and scores higher. Sonnet 5.5 at high, meanwhile, sits within a point of GPT-6 Sol at max for effectively the same cost, which is the comparison most UK teams will care about.

Anthropic’s own charts say the same thing in its own words: Sonnet 5.5 complements Opus 5.5 best at lower effort settings, and at Low or Medium it beats Sonnet 5’s best score on several benchmarks for about a tenth of the cost per task. Its FrontierCode footnote even records Sonnet 5.5 scoring lower at Max than at Xhigh, because at Max it more often ran a multi-agent code-review skill that overran the task.

Two caveats. Artificial Analysis ran every one of these evaluations on a pre-release deployment that carried the structured-outputs bug Anthropic has since fixed, and says it will re-run the relevant ones, so check its live pages before you rely on the decimals. And Anthropic’s own prompting advice is to start at high, the API default, unless the workload is agentic or latency-sensitive; our medium-first recommendation is the cheaper starting point for the routine marketing jobs this post is about. For harder work, take Anthropic’s advice and start at high.

One migration detail matters here too. Anthropic’s what’s new page says the effort levels are recalibrated, so a setting does not produce the same amount of thinking as it did on Sonnet 5. Re-run your effort sweep rather than carrying a value over.

How does Sonnet 5.5 compare with Sonnet 5 and Opus 5.5?

Anthropic’s launch figures show a large jump over Sonnet 5 and near-parity with Opus 5.5 on several evaluations. They are vendor numbers, and Anthropic itself adds that in its own and external testing Opus 5.5 remains clearly stronger at complex, open-ended work. Both statements are true at once.

Anthropic’s reported evaluation results (vendor figures, launch page, checked 29 September 2026)

If the table extends beyond the screen, scroll sideways to view all columns.

Evaluation Sonnet 5.5 Sonnet 5 Opus 5.5
Terminal-Bench 4.0 70.6% 10.3% 66.4% (xhigh)
FrontierCode 1.1, main set 52.1% (xhigh), 46.2% (max) 42.4% 54.4%
CursorBench 4.0 55.5% 34.1% 57.8%
GDPval-AA v2.1 1844 Elo 1449 Elo 1846 Elo
AA-Briefcase v1.1 1811 Elo 1359 Elo 1822 Elo
Humanity’s Last Exam, with tools 64.5% 54.9% 67.7%
OSWorld 2.1, partial 80.1% 57.0% 81.8%
Chartography, no tools 61.6% 15.6% 64.4%

Two caveats come from Anthropic’s own footnotes. The GDPval-AA and AA-Briefcase scores were run by Artificial Analysis on a pre-release deployment with a bug that could degrade structured outputs, since fixed, which Anthropic expects to have understated Sonnet 5.5 slightly if it had any effect. And Anthropic footnotes the Opus 5.5 Terminal-Bench figure as its xhigh score, its best, without saying which setting the Sonnet 5.5 figures were run at.

Artificial Analysis’s independent runs are more mixed than the vendor table. On its Terminal-Bench 4.0 run Sonnet 5.5 scored 64 per cent against 60 per cent for Opus 5.5 and GPT-6 Astra, and it matched Opus 5.5 on AA-Briefcase, GDPval-AA and AutomationBench-AA, with the heavy token use noted above. On factual knowledge it lags: 54 per cent accuracy on AA-Omniscience against 66 per cent for Opus 5.5, though with a lower hallucination rate of 47 per cent against 59. It also sits about six points behind Opus 5.5 on Humanity’s Last Exam and SciCode. For a marketing team, that pattern suggests Sonnet 5.5 for drafting, formatting and tool-driven work, and Opus 5.5 where the facts have to be right first time.

Anthropic also says Sonnet 5.5 is the first Sonnet to finish Pokémon Red working only from screenshots, and that two experts judged its first draft of a 10-slide operating review, built from a public company’s earnings materials and a slide template, ready to send. Treat both as a test list for your own trial, not a forecast.

What breaks if you migrate from Sonnet 5?

Chat users change the model picker and carry on. API integrations need a proper check, because Anthropic’s migration guide lists what changes for code coming from Sonnet 5:

  1. Thinking cannot be disabled. thinking: {"type": "disabled"} is rejected. The lowest setting is now between_tools, which turns off up-front thinking but still returns progress notes between tool calls as thinking blocks. It works at low, medium and high effort only; at xhigh or max you must use adaptive thinking.
  2. Forced tool use returns an error. tool_choice values of any or a named tool are rejected. Use auto with strict tools and a clear instruction in the prompt.
  3. The old computer-use tool is rejected on the Claude API and Google Cloud. Move computer_20251124 to computer_toolset_20260801. Amazon Bedrock keeps the old tool.
  4. Thinking blocks are tied to the model, the conversation and the account. Sonnet 5.5 reads thinking carried over from Sonnet 5, Opus 4.8 and Haiku 4.5, but not from Opus 5, Opus 5.5 or Fable, and a request that routes back down to Haiku 4.5 drops Sonnet 5.5’s blocks. For API accounts created on or after 31 August 2026, editing the system prompt, tools or earlier messages after a Sonnet 5.5 thinking block makes the request fail by default. Keep conversations append-only.
  5. Advisor pairings change. As the executor, Sonnet 5.5 accepts advice only from Mythos, Fable, Opus 5.5, Opus 5 or itself. Opus 4.8, Opus 4.7 and Sonnet 5 advisors return an error.
  6. Longer notes between tool calls now arrive in thinking blocks. Anything beyond a sentence or two moves out of the text stream, and those blocks are empty at the default display setting, so agent progress updates go quiet unless you change it.

Non-default temperature, top_p or top_k values, thinking budgets and assistant prefill also return errors, but they already did on Sonnet 5, so code that has been running there is clear of them. Anthropic’s Claude Code skill can apply most of the changes above across a code base with /claude-api migrate this project to claude-sonnet-5-5. It still produces a checklist for you to verify by hand, and the effort recalibration is the item most likely to change your bill.

Safeguards and fallbacks

Sonnet 5.5’s cyber capabilities are, in Anthropic’s words, comparable to Opus 5’s, so it is the first Sonnet to launch with the cyber safeguards and fallbacks built for the larger models. Routine software work is unaffected, but higher-risk cybersecurity requests visibly fall back to Sonnet 5. Artificial Analysis saw that fallback fire in about 0.1 per cent of its index tasks, mostly in Terminal-Bench, and fall back to Sonnet 5 every time. Biology safeguards are the same as Sonnet 5’s. Anthropic says its expanded Cyber Verification Program will open to vetted defenders soon, and the Life Sciences Verification Program is already open to organisations.

Sonnet 5.5 is also the first Sonnet to ship with classifiers that block reasoning extraction, Anthropic’s response to distillation attacks that use thousands of fake accounts to copy a model’s capabilities. Most developers will not notice. If you move conversations between accounts, including switching accounts mid-session in Claude Code, read the preserved-thinking notes in the docs first. The system card has the full alignment and safeguards detail.

What should UK marketing teams test first?

Pick the everyday workflow that runs on Sonnet 5 today and give Sonnet 5.5 the same brief, the same files and medium effort. Record tokens in, tokens out, tool calls, minutes to finish and minutes of human rework. Then run the same job at high and compare. Do not start at max.

Good first tests for a Manchester or Leeds team:

  • a monthly report deck built from exported data and your own slide template, because Anthropic is making a specific claim about slides
  • a support or sales inbox triage that already runs on Sonnet 5, where Zendesk’s 20 per cent faster ticket handling against its production Claude models is the figure to beat
  • a site audit that rereads the same brand rules on every turn, where fewer tool calls and cache reads at $0.20 decide the cost
  • a fact-heavy brief that must cite primary pages, to see where the AA-Omniscience gap shows up in your work

Treat what comes back as a draft from a capable new starter. Review it, check the citations, and keep a named person accountable for anything that publishes.

Should your business switch to Claude Sonnet 5.5?

For work already on Sonnet 5, yes, after a measured trial, and sooner than you would for an Opus upgrade. The price is unchanged, the speed and token savings are the kind that show up on an invoice, and Sonnet 5 stays available until at least June 2027 if something breaks.

Keep effort at medium or high. Move a job to Opus 5.5 when it needs sustained judgement or has to be right on the facts, and keep Fable 5.1 for the long, expensive runs that Opus 5.5 fails. If GPT-6 Sol or GPT-6 Astra is already in the stack, run it on the same task at the same cost per task. No model wins every job.

If your team wants a measured trial plan for Sonnet 5.5, talk to AIWIZ.

Frequently asked questions

What is Claude Sonnet 5.5?
Anthropic's mid-tier Claude model, released on 28 September 2026 as the second model in the Claude 5.5 family. The API ID is claude-sonnet-5-5, the context window is one million tokens, and Anthropic says it generates output more than 30 per cent faster than Sonnet 5 and costs up to 30 per cent less per task.
How much does Claude Sonnet 5.5 cost?
$2 per million input tokens and $10 per million output, the same as Sonnet 5. Cache reads cost $0.20, five-minute cache writes $2.50, one-hour cache writes $4, and the Batch API halves the input and output rates. Prices checked against Anthropic's pricing page on 29 September 2026.
Is Claude Sonnet 5.5 better than Claude Opus 5.5?
Not overall. On the Artificial Analysis Intelligence Index v4.3.2 it scores 56 at maximum effort against 58 for Opus 5.5, and Anthropic says Opus 5.5 remains clearly stronger at complex, open-ended work. Sonnet 5.5 wins on price at medium and high effort; at maximum effort it costs more per task than Opus 5.5.
Which effort setting should I use with Claude Sonnet 5.5?
Medium for routine work and high where the result needs checking. On the Artificial Analysis index, high effort beats Sonnet 5's best score for about a fifth of the cost per task, while maximum effort costs $7.60 per task, more than Opus 5.5 at the same setting. Effort levels are recalibrated from Sonnet 5, so re-run your sweep.
What breaks when I migrate from Sonnet 5 to Sonnet 5.5?
Disabled thinking, forced tool use and the old computer_20251124 tool all return 400 errors on Sonnet 5.5. Use between_tools for the lowest thinking setting, tool_choice auto with strict tools, and keep conversations append-only, because thinking blocks are tied to the model, the conversation and the account. Effort levels are recalibrated, so re-run your sweep.
Can UK businesses use Claude Sonnet 5.5?
Yes. It is available on the Claude Platform, Amazon Bedrock, Google Cloud and Microsoft Foundry, and in the Claude apps and Claude Code, with zero data retention available. Anthropic bills in US dollars, so convert at your own rate, add VAT where it applies, and check processor terms under UK GDPR before client data goes through it.
Explore
Drag