Anthropic launches Claude Opus 5 with near-flagship benchmarks and relaxed safety guardrails
TECH

Anthropic launches Claude Opus 5 with near-flagship benchmarks and relaxed safety guardrails

28+
Signals

Strategic Overview

  • 01.
    Anthropic released Claude Opus 5 on July 24, 2026, now the default model on Claude Max and the strongest model available on Claude Pro.
  • 02.
    Opus 5 is priced at $5 per million input tokens and $25 per million output tokens, unchanged from Opus 4.8, while Anthropic says it approaches Fable 5-level performance on many tasks.
  • 03.
    The model introduces a low/medium/high effort toggle for cost-versus-capability tradeoffs, and lets developers change available tools mid-conversation without invalidating the prompt cache.
  • 04.
    Opus 5 arrived roughly two months after Opus 4.8 and slots into a model family that also picked up Mythos 5, Fable 5, and Sonnet 5 in June 2026, with Haiku 5 still pending.
  • 05.
    A new Automatic Fallbacks beta feature routes requests flagged by safety classifiers to a less powerful model instead of returning an error, and overall classifier engagement is expected to drop about 85% versus Fable 5.
  • 06.
    Opus 5's full system prompt - about 1,511 lines and 34,000 tokens - was leaked to GitHub within 24 hours of launch, exposing a tool manual, legal-compliance handbook, and commercial-promotion instructions.

Deep Analysis

Near-Flagship Power at Half the Price

Claude Opus 5 launched on July 24, 2026 as the default model on Claude Max and the strongest model available on Claude Pro [1]. Anthropic priced it at $5 per million input tokens and $25 per million output tokens - unchanged from its predecessor Opus 4.8 - while positioning it as approaching the capability of the company's own top-tier Fable 5 model across many tasks [2]. That price-performance claim held up under independent scrutiny: the ARC Prize Foundation confirmed Opus 5 (High reasoning effort) as the highest-scoring model ever recorded on ARC-AGI-3 at 30.2%, clearing five Public Demo environments no prior model had solved, and its verification separately credited Opus 5 with a 42-of-42 gold-medal result across all six IMO 2026 problems, graded by a three-model judge panel plus human experts [3]. The same verification put Opus 5 at 97.5% on the original ARC-AGI-1 benchmark, rounding out a strong reasoning profile [3]. On a head-to-head benchmark comparison, Opus 5 beat Fable 5 outright on 8 of 13 tested benchmarks, including a 9.7-point edge on Frontier-Bench and an 8.5-point edge on AutomationBench [4]. The release also ships a low/medium/high effort toggle so users can trade cost for reasoning depth on a per-task basis, and developers can now swap available tools mid-conversation without invalidating the prompt cache [2].

Anthropic Loosens the Leash on Safety Classifiers

Alongside the benchmark gains, Anthropic re-tuned Opus 5's safety classifiers to trigger roughly 85% less often than they did for Fable 5, routing some flagged requests to a weaker model through a new Automatic Fallbacks beta feature instead of returning an outright refusal [5]. The clearest driver of that shift shows up in the system card: cyber-safety classifier engagement fell from 42% to 5% on FrontierBench, and Opus 5 now permits vulnerability discovery in source code while still blocking binary-based vulnerability scanning and exploit generation [6]. Hugging Face CEO Clem Delangue supplied the industry context for why this mattered, arguing that closed-model APIs routinely flag and refuse legitimate security work because defensive analysis resembles attack preparation [7]. Independent analyst Zvi Mowshowitz pushed back on reading fewer refusals as proof of better alignment, warning that conflating benchmark scores with genuine alignment progress would be a serious mistake for Anthropic to make [6]. The safety picture wasn't purely about loosening restrictions, though: prompt-injection resistance measurably improved too, with a 2.0% attacker success rate versus 5.5% for Opus 4.8 [6].

A 1,511-Line System Prompt Leaks Within a Day

Within 24 hours of launch, a developer published what appears to be Opus 5's complete system prompt to GitHub - roughly 1,511 lines, about 34,000 tokens, 135,027 characters and 19,370 English words [8]. The document reportedly spans a tool and product manual, a legal-compliance handbook, and instructions governing commercial promotion, with the full text preserved in a public system-prompt-leaks repository on GitHub [9]. The exposure matters beyond simple curiosity: coverage of the leak notes that community stress tests reproduced output quality approaching Fable 5 on other, cheaper base models once the leaked prompt was applied - suggesting a meaningful share of what makes Opus 5 feel capable in practice sits in prompt engineering that is no longer proprietary to Anthropic [8].

Chart-Topping Benchmarks, Frustrating First Week

The gap between Opus 5's benchmark sheet and its day-to-day feel emerged fast. Every's Dan Shipper and Katie Parrott spent a week building with the model and called it a hard model to love, reporting that it argued with instructions, stopped before work was finished, and clashed with their existing skills and plugins [10]. That friction sits awkwardly next to enterprise testimony pointing the other way: Cognition CEO Scott Wu praised Opus 5's cost-performance for debugging and root-cause analysis inside Devin, and Zapier CEO Wade Foster said it topped the company's internal automation leaderboard without adding token cost [11]. Both accounts can be true at once - a model can win on the task-completion metrics enterprises track on leaderboards while still frustrating individual developers accustomed to more compliant tooling, a split that recurs across recent flagship launches and is worth watching as Opus 5 sees wider adoption.

Impressive Rankings, Uneasy Community

The social conversation around Opus 5's launch ran hotter and more skeptical than the press coverage. On Reddit, commenters on Anthropic's own announcement flagged what looked like a benchmark chart visually rendering a smaller percentage as better than a larger one, fueling 'benchmaxxing' accusations pending independent verification - a complaint about how the launch numbers were presented, not about the underlying figures. Independent trackers pushed their own rankings into the mix: on YouTube, WorldofAI placed Opus 5 at the top of its own reasoning leaderboard and cited third-party intelligence-index data putting it slightly ahead of Fable 5 and GPT-5.6 Soul at meaningfully lower cost. A separate Reddit thread carried the more eye-catching claim that Opus 5 at 'Low' effort beat Sonnet 5 at 'High' effort for less money, and looser guardrails - but the same thread carried pointed pushback, including reports of elevated hallucination rates and regressions relative to Opus 4.8. That tension between looser guardrails and unintended side effects surfaced elsewhere too: one Reddit report described Opus 5's safety classifiers blocking a legitimate internal security review, prompting the poster to consider open-weight alternatives, while a self-identified Anthropic security researcher countered that Claude models still lead on offensive-security benchmarks. Community sentiment overall read as impressed but wary - strong independent rankings sitting alongside real doubts about benchmark presentation, hallucination behavior, and how well-calibrated the classifier loosening actually is.

Historical Context

2026-05-28
Anthropic released Claude Opus 4.8, the direct predecessor to Opus 5.
2026-06
Anthropic released Mythos 5, Fable 5, and Sonnet 5, establishing the model family that Opus 5 now slots into as the cost-efficient tier.
2026-07-09
OpenAI released GPT-5.6, emphasizing economical token usage ahead of Anthropic's Opus 5 launch.
2026-07-24
Claude Opus 5 officially launched across all Claude platforms alongside an accompanying system card detailing safety-classifier changes.

Power Map

Key Players
Subject

Anthropic launches Claude Opus 5 with near-flagship benchmarks and relaxed safety guardrails

AN

Anthropic

Developer and publisher of Claude Opus 5; positions it as a cheaper, near-flagship alternative to Fable 5 for coding, agentic, and enterprise workflows while pursuing reduced safety-classifier friction for legitimate use.

CO

Cognition (maker of Devin)

Early enterprise adopter; CEO Scott Wu praised Opus 5's cost-performance inside Devin, particularly for debugging and root-cause analysis.

ZA

Zapier

Enterprise adopter; CEO Wade Foster reported Opus 5 topped Zapier's internal automation leaderboard without added token cost.

AR

ARC Prize Foundation

Independent benchmark verifier; confirmed and published Claude Opus 5's record-setting 30.2% ARC-AGI-3 score.

OP

OpenAI

Competitor; released GPT-5.6 on July 9, 2026 with similar emphasis on economical token usage, framing the cost-efficiency comparison narrative around Opus 5's launch.

Fact Check

11 cited
  1. [1] Claude Opus 5
  2. [2] Anthropic launches Claude Opus 5, a cheaper AI model for coding agents and enterprise workflows
  3. [3] Claude Opus 5 Results
  4. [4] Anthropic launches Claude Opus 5 with efficiency, safety improvements
  5. [5] Anthropic launches Opus 5
  6. [6] Claude Opus 5: The System Card
  7. [7] Anthropic debuts Claude Opus 5 with feature that lets users toggle between cost and capability
  8. [8] Claude Opus 5 system prompt leaked
  9. [9] Claude Opus 5 system prompt leak
  10. [10] Vibe Check: Opus 5
  11. [11] Claude Opus 5 HN Analysis

Source Articles

Top 5

THE SIGNAL.

Analysts

"Found Opus 5 powerful but difficult to work with in practice, describing friction with instructions and existing tooling during a week of hands-on use."

Dan Shipper and Katie Parrott
Reviewers, Every (Vibe Check)

"Cautioned against overstating Opus 5's alignment based on reduced classifier engagement alone, while acknowledging real gains in prompt-injection resistance."

Zvi Mowshowitz
Independent AI analyst, Substack (thezvi)

"Argued closed-model APIs over-flag legitimate security research because defensive analysis resembles attack preparation, providing context for why Anthropic loosened Opus 5's cyber classifiers."

Clem Delangue
CEO, Hugging Face
The Crowd

"Anthropic 说他们为 Opus 5 删掉了 Claude Code 80% 的系统提示词。我想验证下是不是。于是把 CLI 指向本地一个小服务器,抓它真正发出去的东西。第一组数字就不太对。Opus 4.7:15,225 字符。Opus 4.8:4,467。Opus 5:7,694。他们是在 4.8 上删的,不是 5。"

@@chenchengpro412

"Claude Opus 5 just got crushed on speed by a model that cost under $1. Same prompt. Same React stack. Fresh terminals. Fable 5: $0.98, 310 seconds; Kimi K3: $0.78, 1,512 seconds; Opus 5: $2.21, 1,529 seconds. Fable 5 was 5x faster than Opus 5 and Kimi K3. Kimi K3 wins on raw..."

@@noclipepe128

"CLAUDE OPUS 5 IS THE WEIRDEST AI RELEASE OF 2026. it beats Fable 5 on the benchmark built to test "can this model learn a new skill on the fly?" and then casually says there's a 41% chance it should be treated as a moral patient."

@@s1rozha_35

"Introducing Claude Opus 5"

@u/ClaudeOfficial2900
Broadcast
Claude Opus 5 Is INSANE – Is This the BEST Model Yet?

Claude Opus 5 Is INSANE – Is This the BEST Model Yet?

Claude Opus 5 Is THE GREATEST AI Model EVER?! Beats Fable & CHEAPER! (Fully Tested)

Claude Opus 5 Is THE GREATEST AI Model EVER?! Beats Fable & CHEAPER! (Fully Tested)

We Tested Claude Opus 5. It's Frustrating with Flashes of Brilliance.

We Tested Claude Opus 5. It's Frustrating with Flashes of Brilliance.