Claude Opus 5: Same Price as Opus 4.8, Nearly 2× the Performance.
Anthropic shipped Opus 5 on July 24 at the old Opus price: $5 per million words in, $25 out. What actually changed, what the effort setting does to your bill, and when Opus 5 is the right model to reach for instead of Sonnet or Fable.
MJBy Micah Johnson · Biggest Goal8 min readUpdated July 26, 2026
Anthropic released Claude Opus 5 last Friday, July 24. It costs exactly what Opus 4.8 cost, and on the hard tests it roughly doubles Opus 4.8's score. If you're on a Claude Max plan you're already using it, because it became your default the day it shipped. Here's what changed, what the effort setting does to your bill, and when Opus 5 is the right model to reach for (and when Sonnet or Fable still is).
Where Claude Opus 5 sits in the lineup
"Claude" isn't one thing. It's a family of models at different levels of power, speed, and cost, and you pick the one that fits the job. We covered this in more depth in the Fable 5 write-up, so this is the short version.
Haiku
Cheap and fast. Simple work at volume.
Sonnet 5
The everyday workhorse. Strong general performance at a reasonable price.
New default
Opus 5 · the heavy tier
Built for hard problems. Now carries a lot of Fable's capability at Opus pricing.
Fable 5
The most capable model Anthropic sells to the public, priced accordingly.
Claude's tiers, lightest to most powerful
The rule has always been that more powerful is not the same as more appropriate. Bigger models cost more and run slower, so the work is in matching the model to the task. Opus 5 creates a grey area around that rule, because it pulls a lot of Fable's capability down into Opus-level pricing.
What actually changed
Start with the bill, since it didn't move. Opus 5 costs $5 per million words in and $25 per million out, the same as Opus 4.8, which is half what Fable charges on input.
Model
Input / M words
Output / M words
Sonnet 5 (workhorse)
$2
$10
Opus 5 (heavy)
$5
$25
Fable 5 (top tier)
$10
$50
Now the performance. On a hard software engineering test, Opus 5 more than doubled the old Opus score while spending less per task. On a test built around problems the model has never seen before, it scored about three times the next best model on the market. And on Zapier's benchmark for finishing real business tasks start to finish, it took the top spot, with even its lowest effort setting completing more tasks than any other model.
2×
Score on a hard engineering test vs. Opus 4.8
$0
Price increase over Opus 4.8
26%
Fewer tokens for the same legal output quality
Three results from the launch worth knowing
It built its own eyes. Anthropic gave Opus 5 a mechanical drawing and asked it to rebuild the part as a 3D model in code, then deliberately took away its ability to look at the drawing. Opus 5 wrote its own image-processing code to pull the geometry out of the raw pixels, then built the part. It did this repeatedly. No competing model solved it in five tries.
Drawing inRebuild this part as a 3D model, in code
Vision removedIts ability to look at the drawing is taken away mid-task
Part built anywayWrote image-processing code to read the geometry out of raw pixels
The drawing test. No competing model solved it in five tries.
It found the real bug. Given a real defect in a widely used developer tool, Opus 5 traced the actual cause and fixed an edge case the human maintainers had missed. A rival model patched the visible symptom and declared victory.
It ran a churn save unattended. Zapier's CEO said Opus 5 took a raw account-health spreadsheet and ran an entire churn-prevention sequence on its own: flagged the at-risk accounts, alerted the right owner, summarized it for the retention team. Earlier models failed the task outright. Opus 5 went 100%.
Reasoning is what nearly every tester report came back to. Opus 5 "thinks" before it writes, and it catches its own bad reasoning during planning instead of after it has built the wrong thing. One engineer described proposing a design, having Opus 5 push back, insisting anyway, and watching it hold its position, explain which part of the idea was worth keeping, and offer a compromise.
It argues with you during planning, not after it has built the wrong thing.
The effort setting is the part that controls your bill
Like 4.8, Opus 5 has an effort setting you can dial from low all the way up to max. Low is cheap and quick. Max spends real money (read: usage) "thinking." The dial itself isn't the news, since 4.8 had it too. What's new is that the low end finally got good enough to trust.
On 4.8, turning the effort down meant accepting worse work, so most people left it high and quietly ate the cost. Opus 5 breaks that tradeoff:
At its lowest setting, it still finishes more of Zapier's business tasks than any other model at any setting.
A legal team held output quality flat while using 26% fewer tokens than 4.8 at max reasoning.
A trading firm got better answers on its own benchmark using roughly a seventh of the reasoning tokens and under half the latency of 4.8.
Now think about what your team actually runs through AI in a week, or should be running through it. Reformatting a list. Drafting a routine email. Pulling three numbers out of a PDF. None of that needs a model's deepest reasoning, and plenty of teams are paying for it anyway, because nobody ever went back and touched the dial.
LowCheap and quick
MediumBalanced
HighDeep reasoning
MaxSpends real usage
80%Send at low effortReformatting, routine email, pulling a few numbers out of a PDF, tidying notes.
20%High to maxModeling, multi-step analysis, long builds.
How to split the dial: cheap by default, expensive on purpose
Anthropic's own charts show Opus 5 beating every other model at any given cost across effort levels, and on some tasks its floor still beats rivals at their ceiling. There's also a Fast mode that runs about two and a half times quicker for double the price, for the times when the clock costs more than the tokens.
The practical question
Which Claude model should your team be using, and at what effort?
The win isn't the most powerful model. It's matching the right one to the real work, so you're not overpaying for reasoning you won't use or underusing what you already pay for. Our free Cowork Masterclass walks non-technical teams through building a simple, practical setup, step by step.
Free · Self-paced · 26 short lessons
Your default probably changed
Opus 5 is now the default on Claude Max and the strongest option on Pro, so if you're on either plan you may already be using it without knowing. It's live everywhere: the Claude app, Claude Code, Cowork, the API, plus Amazon Bedrock, Google Cloud, and Microsoft Foundry.
For most teams, that changes the answer to "which model do we use?":
Sonnet 5 still makes sense for simple, high-volume work where cost is the whole point.
Opus 5 becomes the reasonable default for real work, especially anything multi-step or analytical.
Fable 5 stays parked for the rare monster job, the huge project or the enormous migration, where its extra ceiling is worth the premium.
Where Opus 5 earns its money
Across very different industries, testers reported the same pattern: the harder and longer the task, the further ahead Opus 5 pulls.
Tester
Result with Opus 5
Finance (hard modeling)
About 9 percentage points more accuracy, a third fewer steps, 60% less time
Legal (first-pass redlines)
Nearly double Opus 4.8, using 26% fewer tokens
Box (data analysis)
11% better
Box (due diligence)
17% better
A trading firm's engineer built a market data feed for a new exchange in one sitting, something previous models couldn't finish even with a detailed plan handed to them, and when Opus 5 found no live feed to test against, it built its own test harness.
On longer work, it holds the thread. One tester handed it a chief-of-staff role over their development systems for a weekend, and it built its own monitoring, ran the machines, and surfaced only the decisions that needed a human. And for anyone making documents for a living, testers called it the clearest jump they'd seen from building to revising things like full decks, with better visual judgment and fewer formatting messes.
Where Opus 5 isn't the answer
It's overkill for simple work at volume. If you route everything through Opus 5 because it's the default, you're paying a premium for reasoning you're not using. Sonnet exists for a reason.
It's not the top of the ladder. On the biggest and longest jobs, Fable still has more headroom. Opus 5 gets close on a lot of tasks, not all of them.
It has walls around cybersecurity and biology work. It can find vulnerabilities in source code but blocks the offensive side, and flagged requests get quietly handed to the older Opus 4.8 instead. Anthropic says these limits trip about 85% less often than Fable's do, but if your work lives near those areas you'll feel it occasionally.
The short version
Opus 5 is Anthropic's new daily driver: near top-tier capability, unchanged price, noticeably better judgment, and no data-retention requirement of the kind Fable carries. The one thing to actually do differently is use the effort setting. It's been there since 4.8 and most people still ignore it, but this is the first version where dialing it down doesn't cost you work quality.
Claude Opus 5 is Anthropic's heavy-tier model, released July 24, 2026. It costs the same as the Opus 4.8 it replaces and roughly doubles its score on hard software engineering tests, with noticeably better reasoning and planning. It is now the default model on Claude Max and the strongest option on Pro, and it is available in the Claude app, Claude Code, Cowork, the API, Amazon Bedrock, Google Cloud, and Microsoft Foundry.
How much does Claude Opus 5 cost?
Claude Opus 5 costs $5 per million words of input and $25 per million words of output, exactly what Opus 4.8 cost. That is half what Fable 5 charges on input. There is also a Fast mode that runs about two and a half times quicker for double the price.
What does the effort setting in Claude Opus 5 do?
The effort setting controls how much reasoning Opus 5 spends before it answers, from low up to max, and that spending is what drives your bill. The setting existed on Opus 4.8, but dialing it down meant accepting worse work. On Opus 5 the low end is good enough to trust: at its lowest setting it still finishes more of Zapier's business-task benchmark than any other model at any setting. Send routine work through at low effort and reserve high or max for genuinely hard, multi-step work.
Should I use Claude Opus 5 or Claude Sonnet 5?
Sonnet 5 still makes sense for simple, high-volume work where cost is the whole point. Opus 5 is the reasonable default for real work, especially anything multi-step or analytical. If you route every simple, repetitive task through Opus 5 just because it is the default, you are paying a premium for reasoning you are not using.
Is Claude Opus 5 better than Claude Fable 5?
Opus 5 pulls a lot of Fable 5's capability down to Opus pricing, and it beats Fable on cost at every effort level. Fable 5 still has more headroom on the biggest and longest jobs, like a huge migration or a multi-day project, so keep it parked for those. Opus 5 also trips its cybersecurity and biology safety limits about 85% less often than Fable does, and it carries no 30-day data-retention requirement.
Free Masterclass · $0
A better model won't fix the setup. Build the playbook.
Knowing which model and which effort setting to use is worth real money, and it's only useful once your team has AI wired into actual work. The free Cowork Masterclass walks you through it, step by step, no technical background required.