Anthropic launches Claude Sonnet 5.5
TECH

Anthropic launches Claude Sonnet 5.5

35+
Signals

Strategic Overview

  • 01.
    Anthropic released Claude Sonnet 5.5 on September 28, 2026, positioning it as a faster, cheaper mid-tier companion to Claude Opus 5.5, which had launched about six days earlier.
  • 02.
    Sonnet 5.5 keeps Sonnet 5's pricing - $2 per million input tokens, $10 per million output tokens, and $0.20 per million cache-read tokens - which is half of Opus 5.5's $4/$20 per-million rate.
  • 03.
    On the Terminal-Bench 4.0 agentic coding benchmark, Sonnet 5.5 scored 70.6%, ahead of Opus 5.5's 66.4% and far above Sonnet 5's 10.3%.
  • 04.
    Anthropic says Sonnet 5.5 generates output more than 30% faster than Sonnet 5 and can cut the total cost of completing a task by up to 30%, because it needs fewer tokens and tool calls to get the same result.
  • 05.
    On knowledge-work and computer-use benchmarks - GDPval-AA and OSWorld 2.1 - Sonnet 5.5 scored within one to two points of Opus 5.5, despite the price gap.
  • 06.
    Sonnet 5.5 is the first Sonnet-tier model to launch with the frontier-grade cyber safeguards Anthropic previously reserved for its most capable models, including classifiers meant to block extraction of its own reasoning.
  • 07.
    The model is available on the Claude Platform, Amazon Web Services, Google Cloud, and Microsoft Azure, and has already rolled out inside Cursor's model picker with adjustable effort levels.

The 'cheaper coding king' pitch has an asterisk

Anthropic pitches Claude Sonnet 5.5 as the cheaper coding champion: it scores 70.6% on Terminal-Bench 4.0, beating Opus 5.5's 66.4% and dwarfing Sonnet 5's 10.3%, while costing half as much per token - $2/$10 versus $4/$20 per million [1]. Anthropic and enterprise partners frame this as a straightforward win: up to 30% faster output and up to 30% lower cost per task thanks to fewer tool calls [2]. But that pitch has a caveat. Independent benchmarking firm Artificial Analysis found that at maximum reasoning effort, Sonnet 5.5 wrote roughly 193,000 output tokens per Intelligence Index task - the most it has ever measured for any model, and about 60% more than Opus 5.5 needed for the same tasks [3]. At that effort tier, the 'half the price' math narrows sharply: needing roughly 60% more tokens than Opus 5.5 for the same task erodes much of Sonnet 5.5's per-token discount, even though its rate is still nominally half of Opus 5.5's [4].

The community's real takeaway: Opus as orchestrator, Sonnet as subagent

The most useful discovery to emerge after launch wasn't in Anthropic's own materials - it was a workflow pattern developers converged on almost immediately: use Opus 5.5 as the planner for open-ended reasoning, and delegate well-scoped implementation steps to Sonnet 5.5. That division of labor tracks with how Cursor benchmarks the model - on CursorBench, which grades real coding sessions, Sonnet 5.5's score climbs from 35.8% at low effort to 55.5% at max effort, second only to Opus 5.5, meaning its usefulness depends heavily on how much reasoning budget a task is given [5]. For narrow, well-defined jobs, low or medium effort settings appear to be enough to make Sonnet 5.5 both cheaper and, in several informal head-to-head tests circulating among early adopters, comparable or better in quality than Opus 5.5 run at an equivalently low effort - a direct challenge to any narrative that Sonnet 5.5 is simply a scaled-down Opus. Not every independent test agrees, though: on more open-ended generative tasks, side-by-side comparisons pitting Sonnet 5.5 against Opus 5.5 and rival models have still called Opus the more complete performer, suggesting the coding-benchmark win doesn't automatically carry over to every kind of task.

Frontier-grade safeguards create friction for legitimate security work

Sonnet 5.5 is the first Sonnet-tier model to ship with the frontier-grade cyber safeguards Anthropic previously reserved for its most capable models, including classifiers designed to block extraction of the model's own reasoning [1]. Independent reporting on the model's system card frames this as a defensive measure against misuse of a faster, cheaper coding model [6]. In practice, early adopters report the safeguards can misfire on legitimate work - flagging benign tasks like checking data feeds, changing passwords, or running routine security audits as if they were adversarial probes. That kind of false-positive refusal is a real operational cost for security and IT teams weighing whether to route sensitive workflows through the new model.

Not a full Opus replacement on knowledge-heavy work

Despite the coding and agentic gains, Sonnet 5.5 is not a blanket substitute for Opus 5.5. Artificial Analysis measured it well behind Opus 5.5 on tasks that reward broad knowledge and careful reasoning: 54% versus 66% on factual accuracy, and roughly six points lower on Humanity's Last Exam and SciCode [3]. Anthropic's own benchmarks show the two models much closer on real-world occupational tasks (GDPval-AA) and computer-use (OSWorld 2.1), suggesting the gap is concentrated specifically in fact-recall and scientific-reasoning categories rather than general capability [2].

Historical Context

2026-09-22
Anthropic released Claude Opus 5.5 roughly a week before Sonnet 5.5, its flagship model priced at $4/$20 per million input/output tokens, which Sonnet 5.5 was positioned to complement as a faster, cheaper mid-tier option.
2026-09-28
Sonnet 5.5 launched as a direct upgrade over Sonnet 5, keeping the same per-token pricing while running over 30% faster and scoring far higher on Terminal-Bench 4.0 (70.6% versus Sonnet 5's 10.3%).

Power Map

Key Players
Subject

Anthropic launches Claude Sonnet 5.5

AN

Anthropic

Developer and publisher of Claude Sonnet 5.5 and the broader Claude 5.5 model family

CU

Cursor

AI coding tool that added Sonnet 5.5 to its model picker with adjustable effort levels and benchmarked it via CursorBench

AR

Artificial Analysis

Independent benchmarking firm that measured Sonnet 5.5's Intelligence Index score, effort-level token usage, and gaps versus Opus 5.5 on factual accuracy

BO

Box

Enterprise customer reporting Sonnet 5.5 was 2.4x faster and used 12% fewer tokens than the prior model in production workflows

ZE

Zendesk

Enterprise customer reporting support tickets processed 20% faster using Sonnet 5.5

LO

Lovable

Coding-agent startup reporting Sonnet 5.5 needed one-third fewer tool calls and roughly half as many shell executions

Fact Check

6 cited
  1. [1] Claude Sonnet 5.5 | Anthropic
  2. [2] Anthropic launches Claude Sonnet 5.5 with 30% cost reduction per task due to faster speeds and fewer tool calls
  3. [3] Claude Sonnet 5.5 | Artificial Analysis
  4. [4] Anthropic Releases Claude Sonnet 5.5
  5. [5] Claude Sonnet 5.5 | Cursor Docs
  6. [6] Sonnet 5.5 ships with frontier-grade cyber safeguards

Source Articles

Top 5

THE SIGNAL.

Analysts

“Reports Sonnet 5.5 was more accurate and faster while consuming fewer tokens than the previous model in Box's workflows”

Yashodha Bhavnani, VP at Box
Positive on efficiency gains

“Says the model's speed directly improved customer support throughput”

Abhinay Kathuria, Director of AI at Zendesk
Positive on speed improvement for customer support

“Highlights reduction in tool-call and shell-execution overhead when using Sonnet 5.5 for coding agents”

Fabian Hedin, CTO of Lovable
Positive on agentic efficiency

“Praises Sonnet 5.5's Terminal-Bench performance as near top-tier, while noting it consumes unusually high output tokens per task at maximum effort, undercutting some cost claims, and flags it trails Opus 5.5 on factual accuracy and scientific reasoning benchmarks”

Artificial Analysis (independent benchmarking firm)
Mixed - strong coding gains but a token-efficiency caveat
The Crowd

“Introducing Claude Sonnet 5.5, the second model in the Claude 5.5 family. It's a clear upgrade over Sonnet 5, runs more than 30% faster, and costs up to 30% less for most work.”

@@claudeai54131

“what the fuck, why is Sonnet 5.5 this good?? gave Sonnet and Opus 5.5 the exact same prompt. Sonnet came back with this masterpiece. all in a SINGLE HTML file. Sonnet won this one, it's not even close”

@@filicroval2102

“Claude Sonnet 5.5 vs. Claude Opus 5.5 vs. GPT-6 Astra on generating an animated 3D chariot race using Three.js. >Sonnet 5.5 is competing with GPT-6 Astra, but Sonnet's lower pricing makes it a better value than Astra. >Opus 5.5 is undisputed king.”

@@AI_Screening126

“Introducing Claude Sonnet 5.5, the second model in the Claude 5.5 family”

@u/ClaudeOfficial2200
Broadcast
Introducing Claude Sonnet 5.5

Introducing Claude Sonnet 5.5

Claude Sonnet 5.5 Is Coming But Something Doesn't Add Up

Claude Sonnet 5.5 Is Coming But Something Doesn't Add Up

Anthropic Just Dropped Claude Sonnet 5.5 (MAJOR UPGRADE)

Anthropic Just Dropped Claude Sonnet 5.5 (MAJOR UPGRADE)