Claude Sonnet 5.5 Is Out: The 10 Things That Matter

Duncan RogoffDuncan Rogoff September 28, 2026 8 min read
Earth seen through the window of a spacecraft, with the title Claude Sonnet 5.5
Anthropic

Claude Sonnet 5.5 Is Out Today

Claude Sonnet 5.5 came out today, September 28, 2026. Same price as Sonnet 5. Over 30% faster. And on a test of real jobs done alone, it went from 10.3% to 70.6%.

It's the second model in the Claude 5.5 family. Opus 5.5 came out September 22. Haiku 5.5 is coming in the next few weeks. I went through Anthropic's announcement and pulled out the 10 things that matter if you build with Claude Code.

Sonnet 5.5 Wins the Race

Sonnet 5.5 writes its answers over 30% faster than Sonnet 5. It also needs far fewer tokens to finish a job, so each task costs up to 30% less.

Testers said it batches steps together, so it takes fewer of them. Balyasny ran it on 2,441 finance tasks. Sonnet 5.5 used about 121,000 tokens per answer. Sonnet 5 used 497,000. Zendesk got its support tickets processed 20% faster.

Speed is easier to feel than to read about. The interactive page runs the old and new Sonnet side by side so you can watch the gap. Open the interactive page.

Sonnet 5.5 Costs the Same as Sonnet 5

The price didn't move. Sonnet 5.5 is $2 per million input tokens, $10 per million output tokens, and $0.20 per million for cache reads. That's the same as Sonnet 5.

Opus 5.5 is $4 in and $20 out, twice the price. So you get a faster, smarter model for the same money, and it uses fewer tokens per job on top of that.

How Much You Have to Babysit Sonnet 5.5

This is the number I care about most. Terminal-Bench 4.0 gives the AI a real job and lets it work alone. It only counts if the result works at the end.

Sonnet 5.5 scored 70.6%. Sonnet 5 scored 10.3%. Opus 5.5 scored 66.4%. Those are each model's best setting, where they cost about the same per job. Even on the default Medium effort, Sonnet 5.5 beats Sonnet 5's best score for less than a tenth of the cost per task.

Here's what those jobs look like:

  • Fix a payments system so low-balance alerts arrive within 5 seconds.
  • Fix a lead sign-up form so leads save correctly with no duplicates.
  • Speed up a slow website under time limits.
  • Audit medical insurance claims.

Companies saw the same thing. Base44 ran 118 real app builds. Sonnet 5.5 needed 3.6 tries per build. Opus 5 needed 7.7. It rarely stopped to ask the user questions. Box said it rechecks data in the source documents and catches errors Sonnet 5 missed. Testers also said it gets up to speed on big existing projects fast and handles tasks that run for hours.

Less babysitting means you hand it the job, go do something else, and come back to work that runs.

How Close Sonnet 5.5 Gets to Opus 5.5

Very close, for half the price. On GDPval-AA, which tests real work from 44 jobs, Sonnet 5.5 scored 1844. Opus 5.5 scored 1846. On OSWorld 2.1, Sonnet scored 80.1% and Opus 81.8%. On CursorBench 4.0, 55.5% vs 57.8%.

Opus still comes out ahead on those. But a 2 point gap at twice the price is why Sonnet 5.5 is the one I'd run most of the day.

Sonnet 5.5 Is Good at Gaming Now

Sonnet 5.5 is the first Sonnet to beat Pokémon Red working only from screenshots. It looked at the screen and played.

Creator Kevin Ngo put it this way: "When Claude Opus 5.5 sets the architecture and general framework for a game, I would feel confident in letting Sonnet 5.5 implement it." That's a good way to split the work. Opus plans, Sonnet builds.

Sonnet 5.5 Is a Design Genius

Anthropic gave Sonnet 5.5 a public company's earnings materials and a slide template. They asked for a 10-slide operating review. Two experts judged the first draft ready to send as is.

It also adds more polish to the screens and pages it builds, and it writes more clearly. Testers called it a better partner to work with. Every's Tyler Nishida summed it up in four words: "Claude Sonnet 5.5 cooks."

What People Built With Sonnet 5.5 on Day One

It came out today, so the builds are still rolling in. See what people made on day one on the interactive Sonnet 5.5 page.

Every Sonnet 5.5 Score

Anthropic's scorecard, September 28, 2026. n/a means no score was listed.

TestSonnet 5.5Sonnet 5Opus 5.5GPT-6 Sol
Terminal-Bench 4.070.6%10.3%66.4%n/a
FrontierCode 1.152.1% (46.2% at Max)42.4%54.4%49.3%
CursorBench 4.055.5%34.1%57.8%n/a
GDPval-AA (real work from 44 jobs)1844144918461487
AA-Briefcase1811135918221483
Humanity's Last Exam64.5%54.9%67.7%n/a
OSWorld 2.180.1%57.0%81.8%n/a
Chartography61.6%15.6%64.4%53.6%

Here's what this table says. Opus 5.5 wins every row except Terminal-Bench. Sonnet 5.5 wins that one, and it's the test closest to handing an AI a real job and walking away.

And here's what each model costs. Prices are per million tokens. In is what you send Claude. Out is what Claude writes back.

Anthropic's pricing, September 28, 2026

ModelInOutCompared to Sonnet 5.5
Claude Fable 5.1$10$505x the price
Claude Opus 5.5$4$202x the price
Claude Sonnet 5.5 (new)$2$10Baseline
Claude Sonnet 5 (old)$2$10Same price
Claude Haiku 4.5$1$5Half the price

Where Opus 5.5 Still Beats Sonnet 5.5

Anthropic says Opus 5.5 is clearly stronger at complex, open-ended work that needs steady judgment over a long stretch. Two more things to know about Sonnet 5.5:

  • Max effort runs for a very long time. Medium is the default for a reason.
  • Higher-risk cybersecurity requests fall back to Sonnet 5, and you'll see it happen. Everyday coding and bug fixing are not affected.

On safety, Anthropic's containment tests found Sonnet 5.5 the least likely of any of their models to probe the limits of its test setup.

Here's how I'd split the work:

Use Sonnet 5.5 forUse Opus 5.5 for
Everyday tasks and bug fixesBig projects from scratch
Docs, slides and spreadsheetsHard problems
Fast back-and-forthPlanning
Lots of tasks on a budget

How to Switch to Sonnet 5.5 Today

  1. In the Claude app, open the model menu and pick Sonnet 5.5.
  2. In Claude Code, type /model and choose Sonnet 5.5.
  3. Leave effort on Medium to start. Effort is how hard Claude thinks before it answers. Medium is the default in Claude Code and the apps.

If you're on Pro, Max or Team, you also got one free usage reset. Use it any time before October 22.

Then give it one real job you'd normally sit and watch. See how far it gets alone. And if you want the visual version, watch the old and new Sonnet race side by side on the interactive page.

Free Claude Code drops, straight to your inbox

Short, practical drops on skills, MCP, agents, prompts, and more. No spam, unsubscribe anytime.

Frequently asked questions

When did Claude Sonnet 5.5 come out?

September 28, 2026. It's the second model in the Claude 5.5 family, after Opus 5.5 on September 22. Haiku 5.5 is coming in the next few weeks.

How much does Claude Sonnet 5.5 cost?

The same as Sonnet 5: $2 per million input tokens, $10 per million output tokens, and $0.20 per million for cache reads. Opus 5.5 is $4 and $20, twice the price.

Is Sonnet 5.5 better than Opus 5.5?

On one test. Sonnet 5.5 scored 70.6% on Terminal-Bench 4.0 and Opus 5.5 scored 66.4%. Opus 5.5 wins every other test on Anthropic's scorecard, and Anthropic says Opus is clearly stronger at complex, open-ended work. Sonnet 5.5 is the better pick for everyday work at half the price.

How do I use Sonnet 5.5 in Claude Code?

Type /model in Claude Code and choose Sonnet 5.5. The default effort is Medium. In the Claude app, pick Sonnet 5.5 from the model menu.

What is the free usage reset?

Pro, Max and Team users got one free usage reset with the Sonnet 5.5 launch. You can apply it any time before October 22.

Last reviewed by Duncan Rogoff on September 28, 2026

Duncan Rogoff

Written by

Duncan Rogoff

Apple · PlayStation · Charles Schwab

Keep reading

Claude CodeWorkflows

Claude Code Computer Use: What I Let Claude Click on My Mac (and What It Will Not)

Claude Code computer use lets Claude open your apps, see your screen, click, type, and drag, the way you would, from inside the desktop app. It is off by default, needs a Pro or Max plan, and every app gets approved per session with a fixed level of control: browsers are view-only, terminals are click-only, everything else is full control. Here is how I switched it on, what it caught on a native build that nothing else could reach, which apps I refuse to approve, and the order Claude tries other tools before it touches your screen.

David Iya 9 min
Read article
Claude CodeWorkflows

Claude Code Browser: How I Let Claude Test Its Own Changes in the Desktop App

The Claude Code browser is the Browser pane inside the desktop app's Code tab. Claude starts your dev server, opens the app in the pane, takes screenshots, inspects the DOM, clicks through forms, and fixes what it finds, and by default it does this after every edit. It runs in a clean profile with none of your logins, which is why there is a second option, the Claude in Chrome extension, for anything that has to happen as you. Here is how the two fit together, how auto-verify behaved on a real build, and the prompts I run in the pane now.

David Iya 9 min
Read article

Ready to build it yourself?

Join Claude Code Club, the #1 community for learning claude code, for $9/month.

← Back to the blog