Frontier intelligence at Opus prices: what the benchmarks actually show
Anthropic's benchmark claims are unusually specific about the cost angle, not just raw capability. On Frontier-Bench v0.1, Opus 5 more than doubles predecessor Opus 4.8's score while costing less per task [1]. Independent benchmarking site MarkTechPost logged the exact numbers behind that claim: 43.3% for Opus 5 versus 18.7% for Opus 4.8 and 33.7% for Fable 5 [2]. On CursorBench 3.2 at maximum effort, Opus 5 lands within half a percentage point of Fable 5's peak score at half the cost per task [1], and on OSWorld 2.0, a computer-use benchmark, it beats Fable 5's best result at roughly a third of the cost [1], hitting 70.57% per MarkTechPost's breakdown [2]. Artificial Analysis's own knowledge-work evaluation, AA-Briefcase, put Opus 5 at a 1,720 Elo rating - 146 points ahead of Fable 5's 1,574 - with an Analytical Quality Elo nearly 300 points higher [3]. The pattern across every named benchmark is the same: Opus 5 doesn't need to beat Fable 5 outright to be the more attractive buy, it just needs to get close while costing a fraction as much, at pricing identical to the 4.8 generation it replaces.



