On 24 July 2026, Anthropic released Claude Opus 5 — and left the price exactly where Opus 4.8 sat. Same US$5 per million input tokens, same US$25 per million output. At roughly RM4.70 to the dollar, that is still about RM23.50 in and RM117.50 out per million tokens. If you stopped reading at the pricing table, you would conclude nothing changed for your AI budget this week.
You would be wrong, and the reason is buried in the customer notes, not the headline: one trading firm reports Opus 5 using roughly a seventh of the reasoning tokens of Opus 4.8 at under half the latency on their workloads. Same sticker price, dramatically less metered thinking per task. For Malaysian businesses running Claude on real workflows, the number that matters was never ringgit per million tokens — it is ringgit per completed, correct task. That number just moved.

What Anthropic actually shipped
The Claude Opus 5 announcement is a dense one. The facts that matter for a business buyer:
- Available now on the Claude API (
claude-opus-5), claude.ai, Claude Code, and Claude Cowork. It is the new default on the Max plan and the strongest model on Pro. - Pricing unchanged from Opus 4.8: US$5 / US$25 per million tokens on Anthropic's pricing page. A new Fast Mode runs about 2.5x quicker at 2x the base price — useful for latency-sensitive work, not for batch jobs.
- Coding roughly doubled. Anthropic reports around 2x Opus 4.8's performance on its frontier coding benchmark, and tool vendors like Cursor and Devin describe it landing just under Fable 5 — Anthropic's top-tier model — at half the cost.
- Agent work got materially better. On OSWorld 2.0, the computer-use benchmark, Opus 5 surpasses Fable 5 at about a third of the cost. Zapier reports it completed their account-health automation workflow end to end at 100%.
- Consistency improved. Several early customers highlight lower variance run to run — the same prompt producing dependable output, not a lottery.
The usual caveat applies and Anthropic's competitors would want it said: most of these numbers are vendor-reported benchmarks and partner quotes at launch. Treat them as a strong signal to run your own evaluation, not as a substitute for one.
The number to care about: cost per completed task
Here is the arithmetic your finance team will actually feel.
An agent workflow — say invoice matching or a support triage queue — spends tokens in three ways: reading context, reasoning, and retrying when the first answer is wrong. Opus 5 attacks the second and third. If reasoning-token usage on your workload drops anywhere near what early adopters report, and fewer runs need a retry because variance is down, your effective cost per resolved case falls even though the price sheet is identical. Box, an early tester, quantifies the quality side: about 8% better than Opus 4.8 overall on their document workloads, 11% better on data analysis, and 17% better on due diligence workflows.
That last figure should interest any Malaysian firm doing credit assessment, vendor onboarding, KYC checks, or M&A support — the document-heavy, expensive-to-get-wrong work where we already argued Opus earns its premium. The premium tier just got meaningfully better at precisely that category without costing more.
To put your own numbers on it, our Claude cost calculator estimates monthly spend in ringgit from your team size and workload mix — worth re-running with Opus-tier assumptions if you priced this out earlier in the year.
Does the Opus vs Sonnet line move?
In May we gave a simple rule for choosing between Anthropic's tiers: start from the cost of being wrong, not the benchmark table. That rule survives Opus 5 intact — but the line shifts in one direction.
Work that was borderline — long-running agents, multistep back-office automation, anything where you previously accepted Sonnet's occasional stumble because Opus felt indulgent — now leans Opus. If the model completes an agentic workflow reliably instead of 80% of the way, the human clean-up you avoid is worth more than the token difference. Meanwhile the June release of Claude Sonnet 5 already pulled the mid-tier up, so the honest summary of mid-2026 is: every tier got better, the price sheet stood still, and the argument for keeping senior staff on manual review weakened again.
What has not changed: most business tasks still do not need Opus. FAQ handling, draft content, tagging, templated follow-ups — cheaper models remain the right economics. Reserve Opus 5 for the workflows where errors cost real money or management time.
Where we would point it first in Malaysia and Singapore
Reading the launch through the lens of the 14 agent builds in our use-case directory, three deployments benefit immediately:
- Finance back-office agents — cash application, three-way invoice matching, bank reconciliation. The due-diligence and table-work gains land directly here, and these flows run under PDPA and, for financial institutions, BNM RMiT scrutiny where consistency is a compliance property, not a nicety. This is the core of our AI agent development work, and the tier upgrade changes the economics of it.
- Agent fleets with computer use. If an agent must drive an ERP screen no API reaches — common in Malaysian mid-market stacks — the OSWorld result suggests this is finally practical at sane cost.
- Complex support triage and escalation, where lower variance means the same policy question gets the same answer on Tuesday as it did on Monday. Regulators and customers both notice when it doesn't.
One more practical note: Opus 5 ships with tighter cybersecurity safeguards than its predecessors in some areas — it will assist with vulnerability scanning in source code but blocks penetration-testing and exploit-generation requests, with an enterprise verification programme for legitimate security teams. If you are in security tooling, read the safeguards section of the announcement before committing.
What to do next
Same discipline we recommended for Opus 4.8, updated for the new maths. Do not roll Opus 5 across your stack. Pick the one workflow where being wrong is expensive — proposal review, reconciliation exceptions, due diligence — and measure three things over two weeks: senior review time consumed, correction rate, and effective cost per completed task against whatever model runs it today. If the early-adopter efficiency numbers hold on your workload, the upgrade case writes itself. If they don't, you have spent a fortnight and a few hundred ringgit finding out — which is what a good evaluation costs. Anthropic has also published a prompting guide for Opus 5 on the Claude Platform docs — worth reading before you benchmark, since prompts tuned for older models can understate a new one.
Sources
- Introducing Claude Opus 5 — Anthropic announcement (24 July 2026)
- Anthropic pricing — current per-token API rates
- Claude models overview — Claude Platform docs
MYR figures are indicative at ~RM4.70/USD; check current rates before budgeting.
Konsultasi percuma
Want Opus 5 Evaluated on Your Actual Workflow?
We benchmark Claude models against your real documents and tasks, score them with our Agent GPA evaluation, and give you the cost per resolved case in RM before you commit.

