Claude Sonnet 5.5 Is Out Today
Claude Sonnet 5.5 came out today, September 28, 2026. Same price as Sonnet 5. Over 30% faster. And on a test of real jobs done alone, it went from 10.3% to 70.6%.
It's the second model in the Claude 5.5 family. Opus 5.5 came out September 22. Haiku 5.5 is coming in the next few weeks. I went through Anthropic's announcement and pulled out the 10 things that matter if you build with Claude Code.
Sonnet 5.5 Wins the Race
Sonnet 5.5 writes its answers over 30% faster than Sonnet 5. It also needs far fewer tokens to finish a job, so each task costs up to 30% less.
Testers said it batches steps together, so it takes fewer of them. Balyasny ran it on 2,441 finance tasks. Sonnet 5.5 used about 121,000 tokens per answer. Sonnet 5 used 497,000. Zendesk got its support tickets processed 20% faster.
Speed is easier to feel than to read about. The interactive page runs the old and new Sonnet side by side so you can watch the gap. Open the interactive page.
Sonnet 5.5 Costs the Same as Sonnet 5
The price didn't move. Sonnet 5.5 is $2 per million input tokens, $10 per million output tokens, and $0.20 per million for cache reads. That's the same as Sonnet 5.
Opus 5.5 is $4 in and $20 out, twice the price. So you get a faster, smarter model for the same money, and it uses fewer tokens per job on top of that.
How Much You Have to Babysit Sonnet 5.5
This is the number I care about most. Terminal-Bench 4.0 gives the AI a real job and lets it work alone. It only counts if the result works at the end.
Sonnet 5.5 scored 70.6%. Sonnet 5 scored 10.3%. Opus 5.5 scored 66.4%. Those are each model's best setting, where they cost about the same per job. Even on the default Medium effort, Sonnet 5.5 beats Sonnet 5's best score for less than a tenth of the cost per task.
Here's what those jobs look like:
- Fix a payments system so low-balance alerts arrive within 5 seconds.
- Fix a lead sign-up form so leads save correctly with no duplicates.
- Speed up a slow website under time limits.
- Audit medical insurance claims.
Companies saw the same thing. Base44 ran 118 real app builds. Sonnet 5.5 needed 3.6 tries per build. Opus 5 needed 7.7. It rarely stopped to ask the user questions. Box said it rechecks data in the source documents and catches errors Sonnet 5 missed. Testers also said it gets up to speed on big existing projects fast and handles tasks that run for hours.
Less babysitting means you hand it the job, go do something else, and come back to work that runs.
How Close Sonnet 5.5 Gets to Opus 5.5
Very close, for half the price. On GDPval-AA, which tests real work from 44 jobs, Sonnet 5.5 scored 1844. Opus 5.5 scored 1846. On OSWorld 2.1, Sonnet scored 80.1% and Opus 81.8%. On CursorBench 4.0, 55.5% vs 57.8%.
Opus still comes out ahead on those. But a 2 point gap at twice the price is why Sonnet 5.5 is the one I'd run most of the day.
Sonnet 5.5 Is Good at Gaming Now
Sonnet 5.5 is the first Sonnet to beat Pokémon Red working only from screenshots. It looked at the screen and played.
Creator Kevin Ngo put it this way: "When Claude Opus 5.5 sets the architecture and general framework for a game, I would feel confident in letting Sonnet 5.5 implement it." That's a good way to split the work. Opus plans, Sonnet builds.
Sonnet 5.5 Is a Design Genius
Anthropic gave Sonnet 5.5 a public company's earnings materials and a slide template. They asked for a 10-slide operating review. Two experts judged the first draft ready to send as is.
It also adds more polish to the screens and pages it builds, and it writes more clearly. Testers called it a better partner to work with. Every's Tyler Nishida summed it up in four words: "Claude Sonnet 5.5 cooks."
What People Built With Sonnet 5.5 on Day One
It came out today, so the builds are still rolling in. See what people made on day one on the interactive Sonnet 5.5 page.
Every Sonnet 5.5 Score
Anthropic's scorecard, September 28, 2026. n/a means no score was listed.
| Test | Sonnet 5.5 | Sonnet 5 | Opus 5.5 | GPT-6 Sol |
|---|---|---|---|---|
| Terminal-Bench 4.0 | 70.6% | 10.3% | 66.4% | n/a |
| FrontierCode 1.1 | 52.1% (46.2% at Max) | 42.4% | 54.4% | 49.3% |
| CursorBench 4.0 | 55.5% | 34.1% | 57.8% | n/a |
| GDPval-AA (real work from 44 jobs) | 1844 | 1449 | 1846 | 1487 |
| AA-Briefcase | 1811 | 1359 | 1822 | 1483 |
| Humanity's Last Exam | 64.5% | 54.9% | 67.7% | n/a |
| OSWorld 2.1 | 80.1% | 57.0% | 81.8% | n/a |
| Chartography | 61.6% | 15.6% | 64.4% | 53.6% |
Here's what this table says. Opus 5.5 wins every row except Terminal-Bench. Sonnet 5.5 wins that one, and it's the test closest to handing an AI a real job and walking away.
And here's what each model costs. Prices are per million tokens. In is what you send Claude. Out is what Claude writes back.
Anthropic's pricing, September 28, 2026
| Model | In | Out | Compared to Sonnet 5.5 |
|---|---|---|---|
| Claude Fable 5.1 | $10 | $50 | 5x the price |
| Claude Opus 5.5 | $4 | $20 | 2x the price |
| Claude Sonnet 5.5 (new) | $2 | $10 | Baseline |
| Claude Sonnet 5 (old) | $2 | $10 | Same price |
| Claude Haiku 4.5 | $1 | $5 | Half the price |
Where Opus 5.5 Still Beats Sonnet 5.5
Anthropic says Opus 5.5 is clearly stronger at complex, open-ended work that needs steady judgment over a long stretch. Two more things to know about Sonnet 5.5:
- Max effort runs for a very long time. Medium is the default for a reason.
- Higher-risk cybersecurity requests fall back to Sonnet 5, and you'll see it happen. Everyday coding and bug fixing are not affected.
On safety, Anthropic's containment tests found Sonnet 5.5 the least likely of any of their models to probe the limits of its test setup.
Here's how I'd split the work:
| Use Sonnet 5.5 for | Use Opus 5.5 for |
|---|---|
| Everyday tasks and bug fixes | Big projects from scratch |
| Docs, slides and spreadsheets | Hard problems |
| Fast back-and-forth | Planning |
| Lots of tasks on a budget |
How to Switch to Sonnet 5.5 Today
- In the Claude app, open the model menu and pick Sonnet 5.5.
- In Claude Code, type /model and choose Sonnet 5.5.
- Leave effort on Medium to start. Effort is how hard Claude thinks before it answers. Medium is the default in Claude Code and the apps.
If you're on Pro, Max or Team, you also got one free usage reset. Use it any time before October 22.
Then give it one real job you'd normally sit and watch. See how far it gets alone. And if you want the visual version, watch the old and new Sonnet race side by side on the interactive page.
Short, practical drops on skills, MCP, agents, prompts, and more. No spam, unsubscribe anytime.
Frequently asked questions
When did Claude Sonnet 5.5 come out?
September 28, 2026. It's the second model in the Claude 5.5 family, after Opus 5.5 on September 22. Haiku 5.5 is coming in the next few weeks.
How much does Claude Sonnet 5.5 cost?
The same as Sonnet 5: $2 per million input tokens, $10 per million output tokens, and $0.20 per million for cache reads. Opus 5.5 is $4 and $20, twice the price.
Is Sonnet 5.5 better than Opus 5.5?
On one test. Sonnet 5.5 scored 70.6% on Terminal-Bench 4.0 and Opus 5.5 scored 66.4%. Opus 5.5 wins every other test on Anthropic's scorecard, and Anthropic says Opus is clearly stronger at complex, open-ended work. Sonnet 5.5 is the better pick for everyday work at half the price.
How do I use Sonnet 5.5 in Claude Code?
Type /model in Claude Code and choose Sonnet 5.5. The default effort is Medium. In the Claude app, pick Sonnet 5.5 from the model menu.
What is the free usage reset?
Pro, Max and Team users got one free usage reset with the Sonnet 5.5 launch. You can apply it any time before October 22.
Last reviewed by Duncan Rogoff on September 28, 2026


