Sonnet 5.5: what the cheaper, faster claims mean for your bill
Two write-ups agree on speed and cost. Neither prints the per-token price, so the savings are yours to verify.
By JasonPublished Sep 29, 2026Last verified Sep 29, 20263 min read

The problem
Anthropic released Claude Sonnet 5.5 on September 28, 2026 and pitched it as a faster, cheaper everyday model. If you build with AI coding tools or pay for API usage, the useful question is not whether the launch is exciting but which of the claims will change your own invoice. The two accounts we read agree on the direction (faster, less token use) but emphasize different things, and one of them documents a failure mode at the highest thinking setting. Because the mid-tier model is what most agent workflows default to, a small change in tokens per task compounds quickly. This brief separates what was reported from what still has to be measured on your workload.
Sources
Where this comes from
We didn’t test this ourselves. This is a curated write-up of the reporting below, with our own take added.
Show the 2 sources
- Anthropic releases Sonnet 5.5, which it calls a significantly cheaper, faster work partner
TechCrunch · Sep 28, 2026
We used it for: The release date, the 30% speed claim versus Sonnet 5, reduced token use, the agentic-coding comparison with Opus 5.5, and the planned Haiku follow-up.
- Claude Sonnet 5.5
Simon Willison's Weblog · Sep 28, 2026
We used it for: The same-price-as-Sonnet-5 statement, the up-to-30% cost claim, the author's own thinking-effort runs, and the note that the free claude.ai tier now uses this model.
Our take
Both sources say the model is faster and uses fewer tokens, and one says the price is unchanged from Sonnet 5. Read together, that means any saving arrives through token count, not through a lower rate card, so it will differ from job to job. Neither piece prints a per-million-token price, and we have not verified one, so do not build a budget from the headline. The two also frame the coding story differently: one says Sonnet 5.5 beats Opus 5.5 on agentic coding because you can run several agents for the same money, while the other says it is almost as good as Opus 5.5 on some coding tasks. Those are different claims, and only the second is a like-for-like comparison. The most concrete data point is a warning: the author reports the top thinking setting used 128,000 tokens, cost $1.28 and returned nothing, while a lower setting cost about six cents. Our read is to leave defaults alone until you replay a sample of your own tasks and compare token counts side by side.
What was announced
Anthropic released Claude Sonnet 5.5 on September 28, 2026, roughly three months after Sonnet 5, according to TechCrunch. The outlet reports that it runs about 30% faster than its predecessor, consumes noticeably fewer tokens, and is aimed at everyday work such as coding and document creation. TechCrunch also says a new Haiku model is due in the coming weeks, with no date given.
What the independent write-up adds
Simon Willison states that the price matches Sonnet 5, that it is more than 30% faster, and that most work costs up to 30% less. He also notes that the free tier on claude.ai now runs on Sonnet 5.5.
His own runs are the only hard numbers in either source. At the maximum thinking effort, one image-generation prompt consumed 128,000 tokens, cost $1.28 and produced no result. At the next level down (xhigh), the same kind of prompt took 41 seconds, cost about $0.057 and worked. He reports that Opus 5.5 shows the same maximum-effort problem.

Where the two accounts differ
| Claim | TechCrunch | Willison |
|---|---|---|
| Speed vs Sonnet 5 | About 30% faster | More than 30% faster |
| Cost | Fewer tokens; no price given | Same price; up to 30% cheaper for most work |
| Versus Opus 5.5 | Ahead on agentic coding when several agents run within a budget | Almost as good on some coding tasks |
What is still unknown
No per-token price appears in either source, and neither reports independent benchmark results beyond the vendor's framing. Haiku 5.5 pricing is not yet public.
Best for
- You run API-billed agent or coding jobs on Sonnet 5 todayReplay a sample of real tasks on Sonnet 5.5 and compare token counts before switching defaultsBoth sources describe savings through fewer tokens at an unchanged price, so the gain only shows up in your own usage.
- You set thinking effort to the maximum levelStay one level below (xhigh) until Anthropic addresses the reported problemThe one documented run at maximum effort spent 128,000 tokens for no output, and the same author reports it on Opus 5.5 too.
- You use the free claude.ai tierDo nothing; you are already on the new modelWillison reports the free tier now uses Sonnet 5.5.
Bottom line
Worth a look if you pay per token on Sonnet 5; replay your own jobs before changing any default.
Pick by your situation
| If | Then | Because |
|---|---|---|
| You run API-billed agent or coding jobs on Sonnet 5 today | Replay a sample of real tasks on Sonnet 5.5 and compare token counts before switching defaults | Both sources describe savings through fewer tokens at an unchanged price, so the gain only shows up in your own usage. |
| You set thinking effort to the maximum level | Stay one level below (xhigh) until Anthropic addresses the reported problem | The one documented run at maximum effort spent 128,000 tokens for no output, and the same author reports it on Opus 5.5 too. |
| You use the free claude.ai tier | Do nothing; you are already on the new model | Willison reports the free tier now uses Sonnet 5.5. |
If you run API-billed agent or coding jobs on Sonnet 5 today
Replay a sample of real tasks on Sonnet 5.5 and compare token counts before switching defaults
Both sources describe savings through fewer tokens at an unchanged price, so the gain only shows up in your own usage.
If you set thinking effort to the maximum level
Stay one level below (xhigh) until Anthropic addresses the reported problem
The one documented run at maximum effort spent 128,000 tokens for no output, and the same author reports it on Opus 5.5 too.
If you use the free claude.ai tier
Do nothing; you are already on the new model
Willison reports the free tier now uses Sonnet 5.5.
The direction of the news is credible because two sources agree on speed and lower token use. The size of the saving is not established: no per-token price is published in either piece and the Opus comparison is framed two different ways. Treat the launch as a reason to run a small replay test, not as a reason to migrate everything.
- Confidence
- medium
- Revisit if
- Anthropic publishes Sonnet 5.5 per-token pricing, or the Haiku model ships.
Read next
- Tools & StackShopify lets browser AI agents finish checkout on merchant stores
- Gear & SetupFive monitor placement mistakes and the numbers to fix them