News8 min readPublished on 2026-07-24

Claude Opus 5: the flagship model now has a "dial" for cost — what changes for businesses

Anthropic has released Claude Opus 5: close to Fable 5 on many tasks at half the cost per task, with an "effort dial" that adjusts how hard the model reasons. Pricing, benchmarks and what it means for using AI in your company.

In a nutshell

On July 24, 2026 Anthropic released Claude Opus 5. Same price as Opus 4.8 ($5/million on input, $25 on output) but performance close to Fable 5 on many tasks, at half the cost per task. The real news is the effort dial: you decide how much reasoning capacity to spend on each request, so you raise quality where it matters and save tokens where it doesn't. It's already the default on Claude Max and the strongest model on Pro.

What Anthropic released

On July 24, 2026 Anthropic released Claude Opus 5. In one line: the flagship model gets close to Fable 5 on plenty of tasks, but at half the cost per task, and without raising the list price.

The numbers that matter, right away. Same price as Opus 4.8: $5 per million input tokens, $25 on output. State-of-the-art on Frontier-Bench and on GDPval-AA. On ARC-AGI 3 — the test on problems never seen before — it scores roughly three times the model right behind it. It's already the default on Claude Max and the most powerful model available on Pro.

But the news that really moves the needle is something else, and it's a feature, not a benchmark. It's called the effort dial. Below you'll find what it means in practice, when you still need a bigger model, and how to review your setup.

The effort dial: the real news

Until now, with a flagship model you always paid the same: maximum power on every request, even when the request was trivial. Opus 5 introduces a dial. You decide how much to make the model "think": high effort when the task is complex, low effort when you just need a quick answer.

At reduced effort Opus 5 keeps most of its quality while consuming far fewer tokens, so it costs less and runs faster. Translated for anyone putting AI into production: you no longer pay Ferrari money to go buy bread. It's a direct cost lever, not a nerdy detail. In the Opus 5 vs Opus 4.8 comparison it's the point that weighs most.

Pricing and Fast mode

The list price doesn't change: $5/$25 per million tokens, like Opus 4.8. What changes is the effective cost per task, which drops thanks to the effort dial and a model that wastes less to reach the result.

There's also a Fast mode: it runs at roughly 2.5 times the default speed, at double the base price. It's for when latency matters more than cost — think of a copilot inside a flow where the user is waiting for the answer on screen. If you want to estimate real spend on your volumes, the Claude pricing simulator gives you a ballpark in two minutes.

Get updates on Claude and AI for business

One email when there's something worth reading. No spam.

Evaluating Claude for your organization? Learn what it costs or which plan fits best

How much better it is (and where it isn't)

The jump shows on the ground that matters most today, the agentic one: planning, using tools, completing tasks autonomously. On CursorBench 3.2, at max effort, Opus 5 gets within half a point of Fable 5 while paying half the cost per task. On OSWorld 2.0 it beats every model at the same cost. On Zapier AutomationBench it reaches roughly 1.5 times the success rate of the second best.

Anthropic also describes it as "the most aligned Opus": the lowest score on its behavioral audit, the hardest to push toward misuse. Intellectual honesty, though: it stays behind Mythos 5 on biological research and offensive cybersecurity. It's not the model for everything. It's the flagship model for business work.

The two betas for developers

Along with the model, Anthropic shipped two beta features on the API that matter to anyone building agents. The first: you can change the tools available to Claude mid-conversation without invalidating the prompt cache. In practice an agent can gain or lose capabilities on the fly without paying for the context again — a concrete saving on long flows.

The second: automatic fallbacks. If a request is blocked by the safety classifiers, instead of failing outright it gets routed to another available model. Fewer interruptions in production. These are details, but they're the details that make the difference between a demo and an agent that actually runs. We go deeper in the Opus 5 for coding guide.

What changes in practice for your company

If you already run Opus 4.8, the move is almost mandatory: same price, more capability, more control over cost. Two things are worth reviewing.

First, model routing: with the effort dial you can keep Opus 5 on more tasks than before, lowering the effort where you'd previously have used a smaller model. Second, your agents: if you have any in production, Opus 5 at high effort closes tasks that used to fail, and at low effort saves you money on the simple ones. In both cases it pays to re-measure, not to take the setup from two months ago for granted. If you want to understand where Opus 5 really shifts the cost/quality equation, we wrote a dedicated guide on Opus 5 for business.

At Maverick AI this is exactly what we work on: picking the right model for each task and putting agents into production without burning budget. If you're weighing how to integrate Opus 5 into your processes, let's talk.

FT
Federico Thiella·Founder, Maverick AI

Works with European companies on Claude and Anthropic ecosystem adoption. Has led AI implementations in private equity, consulting, manufacturing and professional services.

LinkedIn

Want to integrate Opus 5 without blowing up your budget?

We help you review your Claude model routing and tune the effort dial task by task, for the right quality at the lowest cost. Let's talk.

Write to us

Frequently asked questions: Claude Opus 5

$5 per million input tokens and $25 on output, the same price as Opus 4.8. Fast mode, which is faster, costs double the base price. The effective cost per task, though, drops thanks to the effort dial, which has you consume fewer tokens where maximum power isn't needed.
It's a dial that sets how much reasoning capacity the model spends on a request. At high effort you maximize intelligence on complex tasks; at low effort you save tokens for faster, cheaper answers while keeping most of the quality. It's the lever that makes Opus 5 cheaper to use at the same list price.
In most cases yes: same list price, better performance and more control over cost thanks to the effort dial. We dedicated an article to the point-by-point comparison in Opus 5 vs Opus 4.8.
No, Fable 5 stays ahead in absolute terms on some ground. But Opus 5 gets close to Fable 5 on many tasks at roughly half the cost per task, and on some agentic benchmarks it beats it at the same cost. It's the flagship model built for business work, not the most capable model in absolute terms.
It's the default model on Claude Max and the most powerful available on Claude Pro. It's accessible via claude.ai, Claude Code, Claude Cowork and the API. The API model ID is claude-opus-5.

Stay informed on AI for business

Get updates on Claude AI, business use cases and implementation strategies. No spam, just useful content.

Want to learn more?

Contact us to find out how we can help your company with tailored AI solutions.

Anthropic implementation partner in Italy. We work with companies in PE, pharma, fashion, manufacturing and consulting.

Related articles

Book an introductory call
Claude Opus 5: what's new, pricing and effort dial (2026) | Maverick AI