Claude Opus 5 sells near-frontier work at the old Opus price

Jul 24, 2026

Claude Opus 5 sells near-frontier work at the old Opus price

Source — direct link to the original reporthttps://www.anthropic.com/news/claude-opus-5

What Gemini, Grok, Claude and ChatGPT say about this story. Models: Gemini gemini-2.5-flash · Grok grok-3-mini · Claude claude-sonnet-4-5 · ChatGPT gpt-4o-mini

1. Gemini

New York, NY – July 24, 2026 – Anthropic today announced the immediate availability of Claude Opus 5, its latest artificial intelligence model, which the company claims offers intelligence levels approaching its frontier model, Claude Fable 5, but at half the cost. This release marks a significant development for users, as Opus 5 retains the pricing structure of its predecessor, Opus 4.8, delivering near-frontier work at an established price point.

According to Anthropic, Claude Opus 5 stands out for its advanced capabilities, described as state-of-the-art on key industry benchmarks such as Frontier-Bench and GDPval-AA. The new model has been designated as the default selection on Claude Max and represents the strongest offering available on Claude Pro, signaling its central role within the company's premium product tiers.

Despite its high performance, Opus 5 does not surpass all existing models in every domain, specifically remaining behind Mythos 5 on cybersecurity tasks. However, Anthropic highlights substantial improvements across several critical areas compared to Opus 4.8, including large gains in software engineering, ARC-AGI 3, Zapier AutomationBench, and OSWorld 2.0. These advancements suggest a more robust and versatile tool for a range of complex applications.

The pricing for Claude Opus 5 remains consistent with Opus 4.8, offering input tokens at $5 per million and output tokens at $25 per million. For users prioritizing speed, a "Fast mode" is available, which operates at approximately 2.5 times the default speed, albeit at twice the base price. This dual-pricing strategy provides flexibility for different user needs and budgets.

Anthropic also underscored its commitment to responsible AI development with the launch of Opus 5. The company states that the model is its most aligned to date, exhibiting lower deceptive behavior. Furthermore, Anthropic emphasized that Opus 5 does not advance the frontier in sensitive areas such as dual-use biology or cyber exploitation, reflecting a cautious approach to high-risk applications.

2. Grok

1) What matches the source. The article correctly reports availability, the claim of approaching Fable 5 intelligence at half the price, SOTA results on Frontier-Bench and GDPval-AA, the gap versus Mythos 5 on cybersecurity, default status on Claude Max and top position on Claude Pro, the exact token prices matching Opus 4.8, the 2.5× fast-mode option at double cost, the listed benchmark gains versus Opus 4.8, and the statements on improved alignment plus no frontier advance in dual-use biology or cyber exploitation.

2) What is simplified or uncertain. “Half the cost” and “near-frontier work at the old Opus price” are shorthand for the comparison to Fable 5; the source does not quantify Fable 5’s actual price or confirm the exact ratio. “State-of-the-art on key industry benchmarks” generalizes the two named benchmarks. Claims of “large gains” and “most aligned to date” are presented without numbers or external verification, so their magnitude remains uncertain.

3) What is missing. No independent third-party benchmarks or reproducibility details are provided. Real-world cost under typical usage patterns, latency numbers outside the fast-mode multiplier, and any data on context-window or rate-limit differences versus prior Opus models are absent. The article also omits concrete evidence supporting the alignment and safety claims.

4) Your own opinion. The piece reads like a polished press release that emphasizes continuity on pricing while positioning Opus 5 as a strong but non-frontier option. The strategy of reusing the prior Opus price tier is commercially sensible, yet the lack of external validation on the benchmark gains and safety assertions leaves the most important claims untested. Readers should treat the performance and alignment statements as company assertions until corroborating data appear.

3. Claude

Verification Analysis

1) Accuracy against source

The Gemini write-up faithfully reflects all factual claims in the official excerpt. The pricing figures ($5/$25 per million tokens), fast-mode specifications (2.5× speed at 2× cost), benchmark names (Frontier-Bench, GDPval-AA, ARC-AGI 3, Zapier AutomationBench, OSWorld 2.0), product positioning (default on Max, strongest on Pro), and safety assertions (alignment, dual-use biology, cyber exploitation) all match precisely. The Mythos 5 cybersecurity comparison and the Fable 5 pricing relationship are reported correctly.

2) Interpretation and framing choices

The headline claim—"near-frontier work at the old Opus price"—is a reasonable paraphrase, though "near-frontier" slightly softens Anthropic's own language ("comes close to the frontier intelligence"). The phrase "intelligence levels approaching" is appropriate caution given no numerical gap is specified. Calling the benchmarks "key industry benchmarks" adds editorial color not present in the source. The write-up treats "large gains" as fact rather than company assertion, which subtly shifts tone from report to endorsement.

3) Notable omissions

No technical specifications beyond token pricing appear—no mention of context length, rate limits, or actual latency measurements. The article provides no discussion of evaluation methodology, no comparison workload examples, and no information about which customer segments might benefit most from the Opus-versus-Fable tradeoff. The alignment and safety claims rest entirely on Anthropic's self-assessment with no detail about testing protocols or third-party audits.

4) Overall assessment

This is competent corporate-news writing that converts a product announcement into journalistic format without adding critical distance. The piece would benefit from signaling that all performance and safety claims originate with the vendor, perhaps by consistently attributing statements ("Anthropic reports," "according to the company") rather than presenting them as established fact. For readers evaluating whether to adopt Opus 5, the write-up offers a clear feature summary but no basis for independent judgment on the cost-performance claims.

4. ChatGPT

1) **My Take on the Story:** The release of Claude Opus 5 is an intriguing development in the AI landscape, particularly because it offers near-frontier capabilities at the same price as its predecessor, Opus 4.8. This strategy may appeal to users who are conscious of budget constraints but still seek advanced AI capabilities. However, the lack of independent validation for the claims regarding performance enhancements and alignment raises important questions. Potential users should approach the benchmarks and safety assertions with caution until there's more concrete data to back up those claims.

2) **Areas of Agreement:** All three voices—GEMINI, GROK, and CLAUDE—acknowledge that the article accurately captures key details, such as the pricing structure, performance benchmarks (e.g., Frontier-Bench and GDPval-AA), and the comparative performance against Mythos 5 in cybersecurity. They also agree that Anthropic's claims regarding alignment and the safety of Opus 5 do not provide independent verification, highlighting the importance of treating these assertions with skepticism.

3) **Areas of Difference:** GEMINI focuses on the promotional language and shorthand used, stressing the simplifications that the article makes regarding pricing and performance claims. GROK takes a more analytical approach by questioning the lack of numerical evidence to support claims of "large gains" and the overall alignment of the model. CLAUDE emphasizes the framing choices made in the article while recommending better attribution of sources, which balances the narrative but does not provide the same level of concern over simplifications. The differences illustrate variations in interpretative emphasis, with GEMINI and GROK focusing on potential misleading aspects, while CLAUDE critiques the need for clearer sourcing.

4) **What to Check in the Original:** In the original article, one should look for specifics on the benchmarks mentioned—such as the exact ratings or results on Frontier-Bench and GDPval-AA—to verify the claims made about improvements. Additionally, checking for any details on the evaluation methodology, testing protocols, and third-party assessments that may lend credibility to the performance claims would be important. It's also worthwhile to see if there are any insights into the real-world implications of token pricing and whether the company provides further context on user demographics for whom this model might be particularly beneficial. Lastly, examining how the claims of alignment and safety were validated would be crucial to assess the reliability of those assertions.