We build with it. Then we review it.
Best & Tested. Hands-on tool reviews · Tested 2026
HomeAI Models › Claude Opus 5 Explained (2026): Benchmarks, Pricing & How It Compares
AI Models Analysis

Claude Opus 5 Explained (2026): Benchmarks, Pricing & How It Compares

Anthropic released Claude Opus 5 on 24 July 2026 and reset the price-to-capability ratio: near-flagship intelligence for the same $5/$25 the old Opus cost. Here's what changed, what the benchmarks actually say, and whether it's the model you should be using.

Claude Opus 5 Explained (2026): Benchmarks, Pricing & How It Compares
On this page (6 sections)
Quick answer

What is Claude Opus 5 and is it worth using? Claude Opus 5 is Anthropic's newest Opus-tier model, released 24 July 2026. It costs $5 per million input tokens and $25 per million output — the same as Opus 4.8 — but Anthropic says it roughly doubles that model's coding score and comes within about 0.5% of the far pricier Fable 5 flagship on coding tasks. For most coding and agent work in 2026, it's the model to use.

Key takeaways
  1. Released 24 July 2026 at $5/$25 per million tokens — unchanged from Opus 4.8.
  2. Anthropic reports it doubles Opus 4.8 on its Frontier-Bench coding benchmark and lands within ~0.5% of Fable 5 on CursorBench at half the cost.
  3. 1M-token context, 128K max output, adaptive thinking, and a new effort toggle for trading cost against capability.
  4. New beta features: mid-conversation tool changes and automatic fallback routing. No data-retention requirement.
  5. It's the new default model on Claude Max and the strongest on Claude Pro.

at a glance

Released24 July 2026
Model IDclaude-opus-5
Pricing$5 / $25 per million tokens (in / out)
Context / output1M tokens in · 128K out
Knowledge cutoffMay 2026
ThinkingAdaptive; effort defaults to high

For most of 2026, the top of Anthropic’s lineup meant paying flagship prices. Claude Opus 5, released on 24 July, changed that. Anthropic’s own line for it is blunt: a model that “comes close to the frontier intelligence of Claude Fable 5 at half the price.” After going through the launch data, that framing mostly holds up.

The important part isn’t that Opus jumped from 4.8 to 5. It’s that the jump came at the same price the old Opus charged. You get a materially stronger model for what you were already paying.

A note on how we cover this. Opus 5 is a day old as we publish. We haven’t run it through our own build brief the way we test app builders, so this is research-grounded analysis, not a hands-on score. Every number below is Anthropic’s own, and it’s dated and linked.

$5 / $25per million tokens (in / out)
1Mtoken context window
May 2026knowledge cutoff

What actually changed

Three things make Opus 5 worth paying attention to.

The price-to-capability reset. Opus 5 costs $5 per million input tokens and $25 per million output. That’s identical to Opus 4.8 and half of Fable 5’s input rate. Anthropic reports it roughly doubles Opus 4.8’s score on Frontier-Bench, its coding benchmark, while charging the same.

The effort toggle. Opus 5 exposes a cost-versus-capability dial — low, medium or high effort per task. Turn it up for hard problems and deep reasoning; turn it down when you want a fast, cheap answer. On the API and in Claude Code it defaults to high. This is the feature the press latched onto, because it hands the cost decision to the user instead of baking it in.

Developer plumbing. Two beta additions: mid-conversation tool changes, so an agent can gain or lose tools without starting a new session, and automatic fallback routing, which lets the API retry on another model when a request needs it. Opus 5 also drops the data-retention requirement that Fable 5 carries, which matters for teams with strict data policies.

The benchmarks, in plain terms

Anthropic published a spread of results at launch. Here’s what they say, translated out of benchmark-speak.

Opus 4.8 on Frontier-Bench coding
~0.5%off Fable 5 on CursorBench (half the cost)
next-best model on ARC-AGI-3

On coding and software work, Opus 5 tops Anthropic’s Frontier-Bench and comes within roughly half a percent of Fable 5 on CursorBench while costing half as much per task. On general knowledge work it scores about three times higher than the next-best model on ARC-AGI-3, and clears Fable 5’s result on the OSWorld computer-use benchmark at a third of the cost. On science, Anthropic measured gains of 10.2 percentage points on organic chemistry and 7.7 on protein prediction over Opus 4.8.

There’s a real limit worth naming. Opus 5 sits behind Anthropic’s restricted Mythos 5 model on biology research and cyber-exploitation tasks, and its safety classifiers are tuned to be about 85% less restrictive than Fable 5’s on narrow cyber work. For ordinary building, none of that touches you. For frontier security or bio research, it does.

Opus 5 vs Opus 4.8 vs Fable 5

Opus 5Opus 4.8Fable 5
Released Jul 2026Jun 2026Jun 2026
Price (in / out) $5 / $25$5 / $25$10 / $50
Coding vs 4.8 ~2× (Frontier-Bench)baselineflagship
Context / output 1M / 128K1M / 128K1M / 128K
Best for Everyday coding & agentsSupersededHardest autonomous work

The takeaway is simple. Against 4.8, there’s no contest — same price, much stronger, so migrate. Against Fable 5, it’s a value call: Opus 5 gives you most of the flagship for half the input cost, and the gap only matters at the extreme end of agentic work.

Who should use Claude Opus 5

USE IT

Coding & agent builders

This is the default in 2026. Strong reasoning over real codebases, a 1M context window for large repos, and the effort toggle to control spend. Migrate here from any older Opus.

See the full lineup →
MAYBE NOT

Pure high-volume / simple jobs

If you are making millions of cheap, simple calls, Sonnet 5 or Haiku 4.5 give you most of the quality for a fraction of the price. Do not pay Opus rates for classification.

Compare the tiers →
STEP UP

The hardest autonomous agents

When the last few percent of capability on long-horizon work pays for itself, Fable 5 is still the flagship. That is the one reason to spend double.

Read the Fable 5 breakdown →

How to build with Opus 5 today

You don’t need to write API code to use it. Opus 5 is live on the Claude API, in Claude Code, and across the chat apps — it’s the default on Claude Max and included with Claude Pro. But if your goal is to ship software, the fastest path is an AI app builder that runs on Claude and points at a current model for you.

Cursor — whose own CursorBench is one of the benchmarks Opus 5 nearly tops — lets you select Claude as the model behind its agent. Prompt-to-app builders like Lovable turn a plain-English brief into a working, deployed app on top of Anthropic’s models, with no setup at all.

The bottom line

Claude Opus 5 is the model most builders should be using right now. It matches the old Opus on price, roughly doubles it on coding, and gets close enough to the Fable 5 flagship that paying double is hard to justify outside the most demanding agent work. If you’re on Opus 4.8, move. If you’re deciding where to start, start here — and if you’re deciding how to put it to work, do it inside a tested app builder rather than from scratch.

Frequently asked questions

How much does Claude Opus 5 cost?
$5 per million input tokens and $25 per million output tokens on the API — identical to Opus 4.8, and half the input price of Fable 5. A fast mode runs at roughly 2× the price for about 2.5× the speed. In the chat apps, it's included with Claude Pro and is the default on Claude Max.
Is Claude Opus 5 better than Opus 4.8?
Yes, and at the same price. Anthropic reports Opus 5 roughly doubles Opus 4.8 on its Frontier-Bench coding benchmark, and posts double-digit gains on scientific tasks — 10.2 percentage points higher on organic chemistry and 7.7 on protein prediction. There's no reason to stay on 4.8 for new work.
Claude Opus 5 vs Fable 5 — which is better?
Fable 5 is still Anthropic's most capable model overall. But Opus 5 is close: within roughly 0.5% on CursorBench and ahead on some agent benchmarks like OSWorld, all at half the input price. Use Fable 5 for the most demanding long-horizon agents; use Opus 5 for essentially everything else.
What is the effort toggle in Claude Opus 5?
Opus 5 lets you set how much effort the model spends on a task — low, medium or high. Higher effort means deeper reasoning and better results at more tokens and latency; lower effort is faster and cheaper. On the API and Claude Code it defaults to high. It's the main dial for trading cost against capability.
What's new in Claude Opus 5 for developers?
Two beta features stand out: mid-conversation tool changes (swap the available tools without restarting a session) and automatic fallback routing (the API can retry on another model when needed). Opus 5 also has no data-retention requirement for general access, unlike Fable 5.
Sources & data
Review changelog

Jul 2026 — Published the day after launch from Anthropic's announcement and specs

M

Marios K.

I build SaaS products and websites for a living. Every tool here is put through the same real-world brief and scored on a fixed rubric — never reviewed from its landing page.

Keep reading