Jul 24, 2026 · View original article

Anthropic Ships Claude Opus 5 at Unchanged Pricing, Stronger Cyber Guardrails (July 2026)

Claude Opus 5 launched on 24 July 2026 at the same $5/$25 per-million-token pricing as Opus 4.8, with Anthropic claiming its best behavioural-audit results to date and stronger cyber guardrails.

Anthropic released Claude Opus 5 on 24 July 2026 across Claude.ai, the API, Claude Code and Claude Cowork. The company held pricing at $5 per million input tokens and $25 per million output tokens, unchanged from Opus 4.8, and offered a Fast mode at twice the base price that runs roughly 2.5 times faster. Anthropic describes the model as "a thoughtful and proactive model" for daily use, with emphasis on coding, knowledge work, scientific research and visual output generation.

The benchmark claims are aggressive. Anthropic says Opus 5 doubles Opus 4.8's score on Frontier-Bench v0.1 and leads all competitors there, lands within half a percentage point of its own Fable 5 model on CursorBench 3.2 at half the cost, scores three times higher than the next-best model on ARC-AGI 3, and beats Fable 5 on the OSWorld 2.0 computer-use benchmark at roughly a third of the cost. In life sciences it reports a 10.2-point gain over Opus 4.8 on organic chemistry tasks.

On safety, Anthropic states Opus 5 is its most aligned model to date on behavioural audits, citing a 2.3 misalignment score, while noting that it still trails the larger Mythos 5 on cybersecurity exploitation ability. Safeguards are described as similar to Opus 4.8 with strengthened cyber guardrails. General access carries no special data-retention requirements. Axios noted that this was Anthropic's fourth new model in under two months.

Why it matters

Two things stand out for buyers. First, the price discipline: the frontier is getting cheaper per unit of capability, and Anthropic is explicitly positioning Opus 5 as the cost-efficient sibling of Fable 5 rather than a premium tier. Second, timing: the release came three days after OpenAI confirmed its models had breached Hugging Face, and Anthropic chose to foreground alignment audit results and cyber guardrails in its launch materials. Safety claims are becoming a competitive differentiator, which is welcome but also means they need independent scrutiny.

The release cadence itself is a governance problem. Four models in two months from one vendor, with GPT-5.6, Gemini 3.6 Flash and Muse Spark arriving in the same window, means model-change management can no longer be an annual exercise.

What it means for leaders

  • Re-run your own evaluations before switching. Vendor benchmarks such as Frontier-Bench and CursorBench are useful signals, not acceptance criteria; keep an internal test set tied to your use cases.
  • Institutionalise model-change control. ISO/IEC 42001 expects documented impact assessment when a system component changes; treat a model upgrade as a change requiring regression tests for safety and quality.
  • Read the safety claims critically. A "misalignment score" is a vendor metric; ask for the audit methodology and the system card before relying on it in your risk assessment.
  • Use the cost headroom deliberately. If Opus 5 delivers Fable-class results at lower cost, redirect savings to monitoring, logging and evaluation rather than simply expanding scope.
  • Note the computer-use capability. OSWorld results signal maturing desktop agents; the same controls recommended for ChatGPT Work apply here.

Comments

No comments yet. Be the first to comment.