Claude Sonnet 5.5 Is Here: 30% Faster at the Same Price — But Do You Still Need Opus 5.5?
Anthropic has released Claude Sonnet 5.5, and the interesting part isn't a lower sticker price. There isn't one. API pricing is identical to Sonnet 5.
What changed is the efficiency. Anthropic says Sonnet 5.5 generates output more than 30% faster than Sonnet 5, and that the total cost of finishing a task can fall by as much as 30% because the model uses fewer tokens and fewer tool calls. Same rate card, smaller bill, at least on Anthropic's telling.
That sets up a question worth sitting with. Anthropic says that in its own tests, Sonnet 5.5 comes strikingly close to Opus 5.5 on several tasks. Opus 5.5 lists at twice the API price. So if the cheaper model gets you most of the way there, when does paying for Opus actually make sense?
This piece doesn't hand you a verdict, because nobody outside Anthropic has one yet. What it does is lay out what changed, what didn't, how the pricing works, and where the honest uncertainty sits.
⚡ Quick facts
- Released: September 28, 2026 — the second model in the Claude 5.5 family
- API input: $2 per million tokens (same as Sonnet 5)
- API output: $10 per million tokens (same as Sonnet 5)
- Cache pricing: $0.20/M for reads, $2.50/M for writes
- Speed claim (Anthropic): output generated 30%+ faster than Sonnet 5
- Task-cost claim (Anthropic): total per-task cost up to 30% lower, thanks to fewer tokens and tool calls
- Opus 5.5 API pricing: $4/M input, $20/M output — twice Sonnet 5.5's listed rates
- Performance evidence: Anthropic's own tests; independent testing hasn't landed yet
- Next: Haiku 5.5 is expected
What actually changed in Claude Sonnet 5.5?
Anthropic released Claude Sonnet 5.5 on Monday, September 28, 2026. It's the second model in the Claude 5.5 family, arriving right after Opus 5.5. Anthropic is positioning it as the workhorse for everyday enterprise work: coding, document creation and spreadsheet tasks are the examples it points to.
If you're looking for a list of shiny new features, this isn't that kind of release. The changes are about how efficiently the model does the job:
- Speed. Anthropic says output is generated more than 30% faster than Sonnet 5.
- Task cost. Anthropic says the total cost of completing a task can drop by as much as 30%.
- The reason for that. The model uses fewer tokens and fewer tool calls to get to an answer.
- Cache pricing. Cache reads are $0.20 per million tokens and cache writes are $2.50 per million, which is cheaper than before.
And here's what didn't change: the headline API price. Sonnet 5.5 costs $2 per million input tokens and $10 per million output tokens, exactly what Sonnet 5 cost.
One note on how to read all of this. The speed and cost figures are Anthropic's numbers. They're a description of what the company saw, not a promise about your workload. Speed in particular depends on what you're asking the model to do, so “more than 30% faster” is best read as a direction, not a guarantee.
The price didn't change. The economics did.
It's easy to skim past the difference between two things that sound the same: the price of a token and the cost of getting something done.
Token price is what's printed on the rate card. $2 per million input tokens, $10 per million output tokens. That number hasn't moved between Sonnet 5 and Sonnet 5.5.
Total cost of a task is what you actually pay at the end. It depends on how many tokens the model reads and writes along the way, and, if it's working as an agent, how many times it calls tools and how much each call adds to the conversation. A model that wanders through a task uses more of everything. A model that goes straight to the answer uses less.
That's the case Anthropic is making for Sonnet 5.5. It says the model needs fewer tokens and fewer tool calls, so the same job costs less to complete, by as much as 30%, even though every token costs exactly what it did before. Anthropic is explicit that this comes from efficiency rather than a lower sticker price.
That's a more important distinction than it sounds. Price cuts are easy to compare. Efficiency gains are harder to see until you run your own workload, and the “up to” in “up to 30%” matters. Your number could be lower. Anthropic hasn't said it will be identical for everyone.
Cheaper cache pricing adds a second angle. If your application reuses large chunks of context, reads from the cache cost $0.20 per million tokens, a tenth of the standard input rate. We don't have a figure for how much that moves the average bill, so it's better to treat it as a lever worth checking than as a number to quote.
It also fits a broader pattern. If you want the bigger picture on how these labs are pricing their models against each other, we covered it in the wider AI pricing battle.
Claude Sonnet 5.5 vs Opus 5.5: what's the real difference?
Here's the part most readers actually came for. On paper, the difference that's easiest to state is the price: at the listed API rates, Opus 5.5 costs twice as much as Sonnet 5.5 per million tokens, on both input and output.
| Claude Sonnet 5.5 | Claude Opus 5.5 | |
|---|---|---|
| API input | $2/M | $4/M |
| API output | $10/M | $20/M |
| Positioning | Everyday workhorse | Higher-end model |
| Release timing | September 28, 2026 | The previous week |
| Performance claim | Close to Opus on several of Anthropic's own tests | The more expensive model in the family |
Those are API prices, meaning what developers and organizations pay per million tokens. They aren't consumer subscription prices, and nothing here tells you what any Claude plan costs.
The last row of the table is where you should slow down. Anthropic says Sonnet 5.5 comes strikingly close to Opus 5.5 on several tasks. That's not the same as saying the two are equal, and it isn't saying Sonnet wins. It's a claim about “several tasks,” in Anthropic's own testing. Which tasks, and how close, matters a lot for whether the claim applies to what you do.
So the honest framing is this. Sonnet 5.5 is the cheaper model, pitched at everyday volume. Opus 5.5 is the pricier one. The gap between them, at least on Anthropic's tests, looks narrower than the price difference would suggest. Whether it stays narrow on real-world workloads is exactly the thing nobody outside Anthropic has been able to check.
If you missed the Opus side of the story, our earlier breakdown of Claude Opus 5.5 covers what that model is about.
Why fewer tokens and fewer tool calls matter
Tokens are easy enough to picture: they're the small chunks of text a model reads and writes, and you're billed for them. Tool calls take a bit more explaining.
When an AI agent works on a task, it often can't finish in one go. It might need to look something up, run a step, check a result, or open a file. Each of those is a tool call. The model asks for the tool, gets something back, reads it, and decides what to do next.
Every one of those round trips has a cost. The model spends tokens deciding to make the call, and then spends more tokens reading what came back. Each step also takes time. String together a long chain of them and even a small task can add up.
Now flip it around. If a model reaches the same result in fewer steps, it does less reading, less writing and less waiting. That's the efficiency Anthropic is describing when it says Sonnet 5.5 uses fewer tokens and fewer tool calls. The saving doesn't come from cheaper tokens. It comes from needing fewer of them.
We don't have a number for how many tool calls Sonnet 5.5 saves on any given task, and it would be a mistake to make one up. What matters is the mechanism: fewer steps means a lower total cost and, usually, a faster finish.
Who should actually care about Sonnet 5.5?
Developers
If you're building on the API, the appeal is straightforward. Faster output can make an application feel more responsive. Fewer tokens and fewer tool calls can lower what each finished task costs. And because the list price is unchanged, moving to the newer model doesn't come with a rate-card surprise. The real question is whether the efficiency shows up in your workload, and the only way to answer that is to try it on your own tasks.
AI agent builders
Agents are where fewer tool calls could matter most. An agent that needs fewer round trips to finish a job spends less on the job and tends to finish sooner. Since agent workloads can involve a lot of back-and-forth, a model that trims the number of steps changes the economics of running one. We won't speculate about how much, because Anthropic hasn't given a figure we can point to.
Enterprises
Efficiency matters more the bigger the volume. A saving that looks small on a single task starts to look large when the same kind of work runs thousands of times a day. That's the audience Anthropic has in mind: it's calling Sonnet 5.5 the workhorse for everyday enterprise work like coding, document creation and spreadsheets.
Casual Claude users
If you just chat with Claude, this pricing conversation mostly isn't about you. The $2 and $10 figures are API prices, which are relevant to developers and organizations building on Claude. Nothing in the announcement details any change to consumer plans, so we won't imply one.
The part we don't know yet
The interesting part is that Anthropic's numbers make Sonnet 5.5 look unusually close to Opus. The less certain part is that these are Anthropic's own tests. Independent testing hasn't landed yet.
That caveat isn't a footnote. Company benchmarks tend to be chosen because they show a model well, and results from a lab's own evaluations don't always carry over to messy real-world work. Independent evaluators have not yet published results on how Sonnet 5.5 holds up, so the following are all still open:
- How close Sonnet 5.5 really gets to Opus 5.5 across a broad range of workloads, not just the tasks Anthropic picked.
- Whether the 30%+ speed improvement holds up outside Anthropic's testing conditions.
- Whether the up-to-30% lower task cost is typical or a best case.
- Where, if anywhere, Opus 5.5 still clearly pulls ahead.
To be clear about what this article is not saying: it's not saying Sonnet 5.5 matches Opus, beats it, or replaces it. It's saying Anthropic claims the two are close on several tasks and that the evidence, for now, comes from Anthropic.
The enterprise angle
Reuters reports that enterprise customers account for roughly 80% of Anthropic's business, and it lists Salesforce, Databricks, Goldman Sachs and Novo Nordisk among them. Those are the supplied examples of the kind of customer base Sonnet 5.5 is meant to serve. It doesn't tell us how any of those companies use Claude, and it shouldn't be read as an endorsement of Sonnet 5.5 from them.
What it does explain is the pitch. When most of your revenue comes from businesses running AI at scale, a model that finishes tasks faster and with fewer tokens is worth more than one that simply tops a leaderboard. Sonnet 5.5's whole positioning, the everyday workhorse for coding, documents and spreadsheets, is built around that logic.
Reuters' headline also ties the launch to Anthropic's planned IPO. That's context, and there's nothing responsible to add about valuation, timing or how investors might react.
The timing: Opus, Sonnet and OpenAI
The sequence is worth spelling out. Opus 5.5 launched the previous week. Sonnet 5.5 followed on September 28. And OpenAI's DevDay is only days away, with Reuters reporting that OpenAI is expected to unveil “o,” described as an always-on agent.
That's the timing, and it's all we'll say about it. Nothing in this reporting says OpenAI has launched “o” yet, and it would be guesswork to predict what it will announce or how it will stack up. If you want the pricing side of the rivalry, that's covered in the article linked above.
A small tension in the timing
One more observation, offered without any drama. Dario Amodei called on the AI industry to slow down releases earlier this month over safety concerns. Anthropic has since shipped two Claude 5.5 models in two weeks.
We're not going to speculate about why, and it would be unfair to call it anything beyond what it is: a contrast between a public statement and a release schedule. Readers can draw their own conclusions. We have our coverage of that call to slow AI development if you want the background.
What's next: Haiku 5.5
Haiku 5.5 is expected next. Anthropic positions it as the fast, high-volume tier, the one for work where speed and scale matter most.
That's everything that's been said. There's no release date, no pricing and no benchmark data for Haiku 5.5, so there's nothing responsible to add beyond “it's coming.”
So, do you still need Opus 5.5?
Depends on what you're doing, and right now the evidence isn't in a state where anyone can give you a universal answer.
What we can say is narrower. Sonnet 5.5 keeps the same API list price as Sonnet 5. Anthropic says it's more than 30% faster and can cut total task costs by up to 30%. Anthropic also says it comes strikingly close to the more expensive Opus 5.5 on several tasks. All of that is Anthropic's own testing, so the real Sonnet-versus-Opus verdict has to wait for independent results. If you run an API workload, the practical move is to test it on your own tasks and see how the numbers hold up.
Frequently asked questions
What is Claude Sonnet 5.5?
Claude Sonnet 5.5 is Anthropic's second Claude 5.5 model, released on September 28, 2026. Anthropic positions it as the workhorse for everyday enterprise tasks such as coding, document creation and spreadsheet work.
Is Claude Sonnet 5.5 cheaper than Opus 5.5?
Yes. At the listed API rates, Sonnet 5.5 costs half as much per million tokens as Opus 5.5. Sonnet 5.5 is $2 per million input tokens and $10 per million output tokens, while Opus 5.5 is $4 per million input tokens and $20 per million output tokens.
Is Sonnet 5.5 as good as Opus 5.5?
Anthropic says Sonnet 5.5 comes strikingly close to Opus 5.5 on several tasks in its own testing. Independent testing hasn't landed yet, so those results shouldn't be treated as proof that the two models perform identically across all workloads.
Is Claude Sonnet 5.5 faster than Sonnet 5?
Anthropic says Sonnet 5.5 generates output more than 30% faster than Sonnet 5. That's Anthropic's figure, and it shouldn't be read as a guarantee for every workload.
Does Sonnet 5.5 cost less than Sonnet 5?
The API sticker price is the same: $2 per million input tokens and $10 per million output tokens. But Anthropic says the total cost of a task can fall by as much as 30% because Sonnet 5.5 uses fewer tokens and fewer tool calls.
Is Haiku 5.5 available?
Not yet. Haiku 5.5 is expected next and is positioned as the fast, high-volume tier. No release date has been given.
Sources
- Reuters — Anthropic rolls out second Claude 5.5 model as it builds toward IPO
- VentureBeat — Anthropic launches Claude Sonnet 5.5 with 30% cost reduction per task
- TestingCatalog — Claude Sonnet 5.5 release date leaks