Claude Sonnet 5.5: what changed, what it costs and how to upgrade

Sonnet 5.5 puts speed and efficient everyday work at the centre of its pitch. Its appeal is the back-and-forth that turns a decent draft into something you want to use.

Claude’s orange starburst logo above Claude Sonnet 5.5 beside folded ivory paper and a terracotta line.
BY
AI EXPERT SYDNEY
PUBLISHED
READ
8 MIN

A decent first draft is useful. An assistant that makes the next six revisions painless is the one you keep using.

That is where Claude Sonnet 5.5 makes its pitch. Anthropic launched it on 28 September 2026, describing a model built for well-scoped coding and everyday work: fixing bugs, making documents, improving slides and working with spreadsheets. The launch announcement reports output generation more than 30% faster than Sonnet 5 and task costs up to 30% lower in Anthropic’s testing, at the same token prices.

That sounds like a more useful kind of upgrade than a model that adds another paragraph explaining how clever it is. If you spend your time making, reviewing and refining things, the quality of the back-and-forth matters.

Sonnet also has a clear place beside the larger model covered in our Claude Opus 5.5 article. The choice is how much thinking the job deserves, and how long you want to wait between changes.

Key points

  • Sonnet 5.5 aims at everyday work you can describe clearly. Coding fixes, documents, slides and spreadsheets are central to Anthropic’s pitch.
  • The speed claim is about output generation. More than 30% faster writing does not make every complete task 30% faster.
  • The base token prices have not fallen. Lower task costs depend on using fewer tokens to reach a useful result.
  • Opus still has a role. Anthropic positions it for more complex, open-ended work requiring sustained judgement.
  • App users and developers have different upgrade paths. Selecting a model is simple; an existing API integration can need compatibility changes.

The real appeal is a shorter gap between idea and revision

Imagine asking for a presentation. You like the outline, but the opening is too long. The chart needs a clearer label. One slide contains a number you do not want rounded. The whole thing would read better with a less formal tone.

Those are small changes, but they decide whether the result feels like your work. A polished first answer can still be wrong for the audience, and the next request is often where you finally explain what you meant.

For me, that is the interesting part of Sonnet’s positioning. A responsive assistant makes it easier to stay involved. You can correct direction while the idea is fresh, rather than waiting through a long answer and trying to remember the next three changes.

The same applies to coding. A developer might know exactly which bug needs fixing and want a capable model to make the change, explain it and respond to feedback. That is a different job from asking an AI to decide how an entire unfamiliar system should be rebuilt.

Speed earns its place when it helps that collaboration. Producing the wrong thing more quickly is still the wrong thing.

Same prices, a different claim about the bill

The phrase “costs less” needs a little unpacking. Anthropic has kept Sonnet 5’s base rates. Its claimed saving comes from needing fewer tokens to do the work.

Tokens are the units of text the model reads and generates. Here is the API comparison with Opus; these are US-dollar usage prices for developers, rather than Claude subscription fees.

API rate, US dollars per million tokens Sonnet 5.5 Opus 5.5
Input US$2 US$4
Output US$10 US$20
Cached input reads US$0.20 US$0.20

Cache writes, tools and some processing arrangements have separate charges. The equal cached-read rate is worth noticing: the “half the price of Opus” comparison applies to ordinary input and output, rather than every part of the bill.

For a deliberately simple hypothetical example, 20,000 billed output tokens cost twenty US cents at Sonnet’s listed output rate. If an equally useful result takes 14,000, that component costs fourteen cents. The rate has stayed the same; less output has reduced the charge by 30%. This example excludes input and other charges and does not describe a result we measured.

The important condition is “equally useful”. A model that gets to the point without wasting steps has improved the economics. One that leaves out a necessary detail has merely made a cheaper mistake.

For someone paying a monthly Claude subscription, API billing is separate, so the claim does not promise a lower renewal price. Better efficiency could make the included usage more useful, but the actual allowance still belongs to the account and plan you have.

When I’d choose Sonnet, and when I’d want more help

The most useful distinction is how settled the job is. Anthropic’s positioning puts Sonnet on well-defined work and Opus on more demanding problems where judgement must continue throughout the task.

Here is how I’d translate that into a starting choice. These are editorial recommendations, rather than results from a head-to-head test.

The job A sensible starting point Why
Fix a bug you can reproduce Sonnet The target is clear and the result can be checked
Turn agreed material into a deck or document Sonnet Most of the work is organising, explaining and refining
Make repeated edits to a design or piece of writing Sonnet Quick responses help you steer the result
Resolve contradictory requirements or plan a major change Compare with Opus The difficult part may be deciding what should happen at all

That last row is where I’d be reluctant to choose on speed alone. If the brief contains a conflict, a fast assistant can confidently polish the wrong solution. Sometimes you need the model to spend longer discovering what the job actually is.

There is no need to turn every everyday task into a contest between models. Start with the one that fits the work, then bring in a stronger option when the problem deserves it.

Where Sonnet 5.5 is available

The release covers the Claude API and several cloud platforms. Anthropic’s availability documentation lists all Claude API customers, Amazon Bedrock, Claude Platform on AWS, Google Cloud and Microsoft Foundry. Provider settings, regions and billing can differ.

Where you use Claude How to reach Sonnet 5.5 Detail to keep in mind
Claude app or web Select it in the model menu when your account has access Your plan determines availability and usage
Claude Code Update and select Sonnet 5.5 explicitly; cloud users may need the provider’s full model ID Sonnet 5.5 requires version 2.1.284 or later
Claude API Use claude-sonnet-5-5 Available to all API customers, with API billing
Supported cloud platforms Select the provider’s Sonnet 5.5 deployment Access and feature support follow that provider’s setup

Anthropic’s Claude Code configuration guide gives the version requirement and explains why the Sonnet alias can select an older model on cloud providers. It also describes the cyber-safety fallback: some higher-risk security requests can run on Sonnet 5 instead, with a visible notice. Routine development remains supported. API access should not be read as a promise of identical access or unlimited usage on every consumer plan.

Claude Haiku 5.5 was also announced as joining the family in the coming weeks. It remains a future release in the Sonnet announcement.

Plenty of capacity, without a reason to make every job bigger

For API use, the model specification lists a one-million-token context window and up to 128,000 output tokens for standard requests. It accepts text and images and produces text. The separate Message Batches API beta supports up to 300,000 output tokens.

Context is what the model can consider; output is what it can produce. The larger batch option is not the limit you should assume for an ordinary interactive answer.

In everyday terms, there is room to work from a substantial set of documents. That is useful when the answer depends on several sources, or when a revision must preserve decisions made earlier in the project.

It does not mean a short editing job benefits from being padded with every file you own. If you want a paragraph rewritten, give it the paragraph, the audience and the reason it needs changing. A model that can do more should still be allowed to do a small job well.

How to upgrade without turning it into a project

For app users, the practical step is to select Sonnet 5.5 where available and give it a familiar task. Something you have edited before makes it easier to notice whether the new model responds well to your direction.

The app’s effort control changes how much reasoning the model uses; higher effort consumes more usage. You do not need to turn it up for every minor rewrite.

Developers have a more substantial caveat. The migration guide documents changes to thinking settings, forced tool calls, conversation history and how progress is returned to an app. Old settings can produce errors, while an interface can appear quiet if it does not display the new progress blocks correctly. Check the integration before moving live traffic.

That is the developer detail worth carrying into a general article. A faster model can still feel unresponsive when the app around it hides the work. The software has to deliver the improvement to the person using it.

The upgrade should feel useful on the fifth change

Sonnet 5.5 has a persuasive everyday role: keep the work moving while you shape it.

I would judge that on the revision loop. Does it preserve the figures when you change the layout? Does it understand which part of a draft you dislike? Does a fix solve the bug without creating another one?

Those are ordinary questions, and they deserve more attention than another impressive first answer. If Sonnet can get through that back-and-forth with less waiting and fewer wasted steps, it could become the model people reach for out of habit.

A good collaborator leaves you with something that feels more like what you meant. That is the promise worth following here.

READING IS FREE. SO IS THE FIRST CONVERSATION.

Want to put this to work in your business? Start there.