The 30 percent pitch, and what's actually behind it

Anthropic's official announcement frames Sonnet 5.5 as the second model in the Claude 5.5 family [1]: it generates output more than 30% faster and cuts the total cost of completing a task by up to 30% versus Sonnet 5 [3]. Notably, that saving is not a price cut - the per-token rates are frozen at Sonnet 5's exact numbers, $2 per million input tokens and $10 per million output tokens [2]- it comes entirely from the model doing the same job with fewer tokens and fewer tool calls [3]. Early enterprise testers back the mechanism with concrete numbers: Box reported working 2.4x faster while using 12% fewer total tokens, Slack cut output tokens by 14%, Lovable used roughly a third fewer tool calls, and Atlassian measured 30% faster agent operations [3]. Zendesk's Director of AI, Abhinay Kathuria, said the model made fewer wrong decisions and resolved tickets faster than the Claude models Zendesk used in production, with tickets processed 20% faster [4].


