Anthropic released Claude Haiku 5.5, its cheapest and fastest small model yet, cutting API prices up to 90% for typical requests, matching OpenAI's GPT-6 Luna on base pricing, and adding Haiku's first adjustable effort dial - all roughly two weeks after Claude Opus 5.5 completed the Claude 5.5 family.
TECH

Anthropic released Claude Haiku 5.5, its cheapest and fastest small model yet, cutting API prices up to 90% for typical requests, matching OpenAI's GPT-6 Luna on base pricing, and adding Haiku's first adjustable effort dial - all roughly two weeks after Claude Opus 5.5 completed the Claude 5.5 family.

25+
Signals

Strategic Overview

  • 01.
    Anthropic released Claude Haiku 5.5 on October 7, 2026, calling it the cheapest, fastest, and most capable small model it has shipped to date, designed for high-volume, cost-sensitive tasks like summaries, compactions, database queries, and classification.
  • 02.
    The model is available immediately on the Claude Platform, Amazon Web Services, Google Cloud, and Microsoft Azure under the model ID claude-haiku-5-5.
  • 03.
    Haiku 5.5 introduces Anthropic's first Haiku-class adjustable effort setting, letting users trade cost for smarter answers across levels including Low, Med, High, Xhigh, and Max.
  • 04.
    Pricing is $0.10 per million input tokens and $0.50 per million output tokens for prompts under 100,000 tokens, rising to $0.50/$2.50 per million for longer prompts, with cache reads at $0.01/$0.05 and cache writes at $0.125/$0.625 per million at the respective tiers.
  • 05.
    Haiku 5.5 costs about 75% less on average to run than Haiku 4.5, and roughly 90% less specifically for requests under 100,000 tokens, which Anthropic says make up about 90% of Haiku traffic.
  • 06.
    Haiku 5.5's launch came 15 days after Anthropic's Opus 5.5 release on September 22, 2026, rounding out the Claude 5.5 model family.
  • 07.
    On Anthropic's reported benchmarks, Haiku 5.5 posts large gains over Haiku 4.5: OSWorld 2.1 at 72.4% versus 15.7%, Humanity's Last Exam (no tools) at 45.9% versus 10.2%, Terminal-Bench 4.0 at 39.2% versus 0.0%, GDPval-AA v2.1 at 1620 versus 735, and AA-Briefcase v1.1 at 1578 versus 614.

The 100,000-Token Cliff Behind Haiku 5.5's Headline Price

Anthropic is advertising Claude Haiku 5.5 as costing $0.10 per million input tokens and $0.50 per million output tokens, a rate that holds only for prompts under 100,000 tokens [1]. Several outlets covering the launch flagged that this threshold functions as a pricing cliff: cross it, and the rate more than quintuples, to $0.50 per million input tokens and $2.50 per million output tokens [1][2].

That design explains the gap between Anthropic's two headline savings figures. The company says Haiku 5.5 costs about 75% less to run than Haiku 4.5 on average, but close to 90% less specifically for requests that stay under the 100,000-token line - which Anthropic estimates covers roughly 90% of actual Haiku traffic [3]. Developers migrating long-context workloads from Haiku 4.5 are the ones most likely to see the smaller end of that discount rather than the splashier number in Anthropic's own announcement.

An Effort Dial, and the Pareto-Frontier Pushback

Haiku 5.5 ships with Anthropic's first Haiku-class "effort" setting, letting developers choose among Low, Medium, High, Xhigh, and Max levels to trade inference cost against answer quality [4]. It is a new control surface for a model line that has historically been single-speed, letting the same model ID serve both the cheapest, fastest classification jobs and harder reasoning work.

Community testing complicated that pitch. Discussion on Reddit pointed to Anthropic's own Terminal-Bench results, arguing Haiku's effort curve falls below the Pareto frontier once effort is turned up to Medium or higher - meaning that for anything but the shortest, cheapest tasks, Sonnet may deliver better results for the money. A second line of community pushback argued that price-per-token understates true cost because Haiku tends to consume more tokens per completed task than larger models, making price-per-completed-task the more honest comparison. Neither critique disputes that Haiku 5.5 is cheap at the low-effort end; they argue the dial does not hold its value as you turn it up.

Price War at the Bottom: Matching GPT-6 Luna

Price War at the Bottom: Matching GPT-6 Luna
Claude Haiku 5.5 vs. Haiku 4.5 and GPT-6 Luna on OSWorld 2.1 and Terminal-Bench 4.0, as reported by Anthropic.

Haiku 5.5's base pricing was set to match OpenAI's GPT-6 Luna exactly, and Anthropic is using benchmark comparisons against Luna - not just its own prior Haiku model - to make its case [3]. On OSWorld 2.1, Anthropic reports Haiku 5.5 at 72.4% versus 48.9% for GPT-6 Luna, and on Terminal-Bench 4.0, 39.2% versus 16.4% [1][3].

Coverage of the launch tied the pricing move to timing beyond the model itself: one outlet's headline situated the release "eight days before the October 15 deadline" [2], while other reporting framed it as arriving roughly a month ahead of Anthropic's planned November IPO [5]. That combination - frontier-lab economics intersecting a public-offering timeline - is part of why some coverage described the release less as a routine model refresh and more as a signal of margin compression at the bottom of the LLM stack, where Anthropic and OpenAI are now competing on cost per token as much as on capability.

Benchmarks Aside, the Real Pitch May Be Usage Efficiency

Anthropic's own benchmark table shows large jumps over Haiku 4.5 - Humanity's Last Exam climbing from 10.2% to 45.9% without tools, GDPval-AA v2.1 more than doubling from 735 to 1620 [1]. Customer references echo that: a staff software engineer quoted by Anthropic reported over a 30% reduction in latency for task completions and up to 2.5x faster inference per agent turn versus the model it replaced, and HubSpot said Haiku 5.5 scored 92.8% on its internal benchmark suite, its best result yet [1].

One widely watched independent test framed the value proposition differently: a full suite of agentic coding tasks run against Haiku 5.5 consumed only about 1% of a creator's weekly Claude.ai usage allowance on the $200/month plan, versus roughly 10% for other models tested that same week - a gap the reviewer called more persuasive than any individual benchmark score. The same test placed Haiku 5.5's raw capability close to state-of-the-art from six months earlier, a meaningful jump after a long stretch with no new Haiku release, while still rating it behind Sonnet, Opus, and GPT-6 Astra on harder tasks.

Historical Context

2024-03-13
First model in the Haiku line, launched as a fast, lower-latency option in the Claude 3 family.
2024-11-04
Faster, lower-cost successor released as part of the Claude 3.5 family.
2025-10-15
Prior-generation Haiku release and the first Haiku model with extended thinking and computer-use capabilities.
2026-09-22
Anthropic released Opus 5.5 roughly two weeks before Haiku 5.5, as part of the broader Claude 5.5 family rollout.
2026-10-07
Released, completing the Claude 5.5 model family and cutting small-model pricing sharply.

Power Map

Key Players
Subject

Anthropic released Claude Haiku 5.5, its cheapest and fastest small model yet, cutting API prices up to 90% for typical requests, matching OpenAI's GPT-6 Luna on base pricing, and adding Haiku's first adjustable effort dial - all roughly two weeks after Claude Opus 5.5 completed the Claude 5.5 family.

AN

Anthropic

Model developer and publisher; positions Haiku 5.5 as a low-cost, high-throughput supporting model beneath Opus and Sonnet in its lineup, ahead of a planned IPO.

OP

OpenAI (GPT-6 Luna)

Primary competitor at the small-model price/performance tier; Haiku 5.5's base pricing matches GPT-6 Luna, and Anthropic claims to lead Luna on benchmarks where both appear, such as OSWorld (72.4% vs 48.9%).

AW

AWS, Google Cloud, Microsoft Azure

Cloud distribution partners making Haiku 5.5 available alongside the Claude Platform at launch.

RO

Rogo (Alex Wang)

Applied-AI customer validating Haiku 5.5's accuracy-for-cost tradeoff in production use.

HU

HubSpot

Customer reporting benchmark results on an internal test suite, citing high accuracy from Haiku 5.5.

GI

GitHub Copilot

Rolled out Haiku 5.5 for quick edits and smaller coding tasks within Copilot.

Fact Check

6 cited
  1. [1] Claude Haiku 5.5
  2. [2] Anthropic Haiku 5.5 Collapses the Small-Model Pricing Floor at $0.10/$0.50, Eight Days Before the October 15 Deadline
  3. [3] Anthropic Launches Claude Haiku 5.5 With 90% API Price Reduction, Matching GPT-6 Luna
  4. [4] Anthropic Launches Haiku 5.5, Its Cheapest, Fastest Claude Model
  5. [5] Anthropic Launches Claude Haiku 5.5 Ahead of November $2T IPO
  6. [6] GitHub Rolls Out Haiku 5.5 in Copilot for Quick Edits and Smaller Tasks

Source Articles

Top 1

THE SIGNAL.

Analysts

“Says Haiku 5.5 is accurate enough to trust for production tasks while being fast and cheap enough to run at high volume.”

Alex Wang
Rogo (applied AI customer)

“Highlights large latency and inference-speed improvements over the previous model used in production.”

Aaron Vinh
Staff Software Engineer (customer reference)

“Reports Haiku 5.5 scored 92.8% averaged over three runs, the best result HubSpot has seen yet on its internal benchmark suite.”

Ze'ev Klapow
Distinguished Software Engineer, HubSpot
The Crowd

“Introducing Claude Haiku 5.5: the cheapest, fastest, and most capable small model we've ever released. On average, it costs around 75% less to run than Claude Haiku 4.5.”

@@claudeai35713

“Haiku 5.5 is now available in the Claude Platform and Claude Code. On average, it costs around 75% less to run than Haiku 4.5. It pairs well with Opus 5.5 or Sonnet 5.5 as a subagent. Use it for high-volume, cost-sensitive tasks like summaries, compactions, or database queries.”

@@ClaudeDevs9457

“Claude Code 2.1.293 is now available. 56 CLI changes Highlights: • Haiku default changed to Claude Haiku 5.5 (1M context), allowing larger prompts and reducing per-token cost • Fixed action-order in context compaction so done work isn't retracted or redone, yielding stable”

@@ClaudeCodeLog218

“Introducing Claude Haiku 5.5: the cheapest, fastest, and most capable small model we've ever released”

@u/ClaudeOfficial1800
Broadcast
Claude Haiku 5.5 Is INSANE – This IS the BEST Cheap Model Yet!

Claude Haiku 5.5 Is INSANE – This IS the BEST Cheap Model Yet!

Claude Haiku 5.5: Anthropic's New Workhorse.

Claude Haiku 5.5: Anthropic's New Workhorse.

Claude Haiku 5.5 - Pricing and Benchmark | Beats GPT-6 Luna? | Best & Cheapest Model

Claude Haiku 5.5 - Pricing and Benchmark | Beats GPT-6 Luna? | Best & Cheapest Model

Anthropic released Claude Haiku 5.5, its cheapest and fastest small model yet, cutting API prices up to 90% for typical requests, matching OpenAI's GPT-6 Luna on base pricing, and adding Haiku's first adjustable effort dial - all roughly two weeks after Claude Opus 5.5 completed the Claude 5.5 family. — AI News | Agentic Brew