Claude Opus 5 Outscores Fable 5 on Most Benchmarks—At Half the Price

Anthropic, a leading artificial intelligence safety and research company, today announced the immediate availability of Claude Opus 5, a new flagship large language model (LLM) designed to deliver superior performance and cost efficiency for businesses. This release marks a significant strategic pivot for Anthropic, as Opus 5 not only comes at a lower price point…

 Avatar

by

11 minutes

Read Time

Anthropic, a leading artificial intelligence safety and research company, today announced the immediate availability of Claude Opus 5, a new flagship large language model (LLM) designed to deliver superior performance and cost efficiency for businesses. This release marks a significant strategic pivot for Anthropic, as Opus 5 not only comes at a lower price point for commercial deployment compared to its previously positioned "everyday frontier product," Claude Fable 5, but also demonstrably surpasses it across a range of critical benchmarks. The launch positions Opus 5 as the new heavy workhorse in Anthropic’s diverse AI model lineup, filling a crucial role that Fable 5, plagued by a tumultuous launch and subsequent restrictions, ultimately could not sustain.

Anthropic’s Evolving Model Hierarchy and Strategic Vision

To fully appreciate the significance of Opus 5, it is essential to understand Anthropic’s structured approach to its AI model offerings. The company maintains a four-tiered hierarchy, each designed to cater to specific user needs and computational demands. At the base is Haiku, optimized for speed and cost-effectiveness, ideal for simpler, high-volume tasks. Next is Sonnet, a mid-range model that strikes a balance between performance and efficiency, suitable for a broader array of general applications. Opus historically represented the company’s most capable and complex model, designed for heavy computational workloads and sophisticated problem-solving. Above this, Anthropic introduced the Mythos class earlier this spring, a new tier intended for the absolute frontier of AI capabilities, often with enhanced safety protocols.

Within the Mythos class, two models were initially positioned: Claude Fable 5, intended for general public access as the subscriber flagship, and Claude Mythos 5, a highly restricted version with fewer guardrails, reserved exclusively for vetted cybersecurity researchers and critical infrastructure operators through a specialized initiative known as Project Glasswing. Project Glasswing underscores Anthropic’s commitment to exploring the full potential of advanced AI in controlled environments, balancing innovation with stringent safety and ethical considerations, particularly in high-stakes domains. The introduction of Opus 5 now redefines the hierarchy, effectively repositioning the "Opus" line as the practical, high-performance standard for the majority of paying users.

The Tumultuous Rollout of Claude Fable 5

The path to Opus 5’s prominence was paved, in part, by the challenging journey of its intended predecessor, Claude Fable 5. Launched with considerable anticipation on June 9, Fable 5 was positioned as Anthropic’s cutting-edge offering for everyday users, integrated into standard subscription plans. However, its tenure as the flagship was remarkably short-lived and fraught with complications. Just three days after its release, on June 12, the U.S. government issued an emergency export control order, compelling Anthropic to globally pull both Claude Fable 5 and Claude Mythos 5 from availability. The reason cited was a critical "jailbreak vulnerability," a term referring to methods used to bypass an AI model’s built-in safety mechanisms and elicit responses that violate its ethical guidelines or intended restrictions.

This incident highlighted the increasing scrutiny and regulatory challenges faced by AI developers, particularly concerning potential dual-use capabilities and the risks of misuse. The swift government intervention underscored the nascent but rapidly evolving landscape of AI governance and national security implications. Anthropic worked diligently to address the vulnerability, and after a period of intense remediation, Fable 5 was brought back online on June 30, following the lifting of the export controls. However, its return was not without significant changes: Fable 5 was immediately shifted to a "credits-only" model, effectively removing it from standard subscription plans and limiting its accessibility. This change severely hampered its intended role as the readily available, everyday frontier product, leaving a conspicuous void in Anthropic’s public-facing top-tier offering. It is into this void that Claude Opus 5 now confidently steps, providing a stable, high-performing, and cost-effective alternative.

Opus 5’s Unprecedented Performance Gains

The core strength of Claude Opus 5 lies in its remarkable performance, which consistently surpasses Fable 5 and often competes favorably, or even outperforms, OpenAI’s leading commercial rival, GPT-5.6 Sol, across a suite of demanding benchmarks. This comprehensive superiority extends beyond mere numerical improvements, encompassing greater reliability and consistency in execution.

One of the most telling evaluations is Frontier-Bench v0.1, a benchmark specifically designed to assess whether AI coding agents can complete complex software engineering tasks from start to finish. Scored as a percentage of tasks successfully passed, Opus 5 achieved an impressive 43.3%. In stark contrast, Fable 5 managed only 33.7%, while OpenAI’s GPT-5.6 Sol scored 34.4%. This significant lead for Opus 5 in agentic coding tasks suggests a substantial leap in its ability to understand, plan, and execute complex programming instructions, a critical capability for developers and enterprises looking to automate software development workflows.

The widest margin of superiority for Opus 5 is observed on ARC-AGI-3, a benchmark focused on genuine problem-solving rather than rote memorization. This test presents novel puzzles that a model could not have encountered during its training, demanding true understanding and reasoning capabilities. Opus 5 achieved a remarkable 30.2% of puzzles solved, dwarfing GPT-5.6 Sol’s 7.8%. Notably, Fable 5 was not tested on ARC-AGI-3, a decision that could be interpreted as an acknowledgment of its limitations in complex reasoning or perhaps a strategic choice given its earlier performance issues. The dramatic difference highlights Opus 5’s advanced cognitive abilities, positioning it as a powerful tool for tasks requiring creative problem-solving and abstract reasoning.

Furthermore, on GDPval-AA v2, a comprehensive knowledge work benchmark that employs Elo ratings – the chess-style ranking system used to measure relative performance on real professional tasks – Opus 5 achieved a score of 1,861. This compares favorably against Fable 5’s 1,747 and GPT-5.6 Sol’s 1,736. Higher Elo ratings in this context signify a model’s superior capability in handling diverse professional tasks, from analysis and synthesis to decision support, making Opus 5 a more effective assistant for knowledge workers across various industries.

Third-Party Endorsements and Real-World Impact

The superior performance of Opus 5 is not merely confined to benchmark scores; it is validated by real-world applications and enthusiastic endorsements from key industry partners and developers.

Lovable, a prominent developer platform serving millions of users, conducted its own internal evaluations of Opus 5. Fabian Hedin, in a statement shared by Anthropic, underscored the qualitative improvements: "It isn’t just better on our hardest agentic coding tasks, up 22% over Opus 4.7, it’s steadier, with far less variance run to run." This emphasis on steadiness and reduced variance is crucial for enterprise adoption, where predictability and reliability are often as important as raw performance. For developers, a model that consistently performs well without unpredictable fluctuations reduces debugging time and increases trust in automated processes.

Claude Opus 5 Outscores Fable 5 on Most Benchmarks—At Half the Price

Zapier, a leading automation platform, tested Opus 5 on its AutomationBench, an evaluation designed to assess a model’s ability to execute full business workflows autonomously. Their findings were unequivocally positive: the model "took a raw account-health workbook and ran a full churn-prevention sequence end to end: flagging at-risk accounts, alerting the right owner, and summarizing for retention ops. Previous models didn’t pass; Opus 5 hit 100%." This demonstration of end-to-end automation capabilities without human intervention is a powerful testament to Opus 5’s potential to revolutionize business operations, enabling more efficient and intelligent workflow management.

Beyond business automation, Anthropic is also actively positioning Opus 5 as a significant upgrade for scientific research. Ultima Genomics, a company specializing in DNA sequencing, provided a compelling testimonial, stating that the model "behaves more like a careful scientist than any model we’ve run. It reaches for the right statistical tests to rule out confounders, cross-checks its own results by independent methods, and stays on track through long multi-step analyses." This highlights Opus 5’s capacity for rigorous scientific methodology, including hypothesis testing, data validation, and maintaining coherence through complex analytical processes – capabilities that could accelerate discovery in fields like bioinformatics, drug development, and materials science.

It is worth noting the singular exceptions where Fable 5 maintains a marginal lead: specific tasks within the legal and health sectors. While the exact reasons for this slight advantage are not fully detailed, it could suggest that Fable 5’s initial training data or safety guardrails were particularly tuned for the highly sensitive and regulated nature of these domains, even if its overall general intelligence and reliability were subsequently found lacking.

Pricing, Accessibility, and the Paradox of the Mythos Class

Opus 5 is now available to developers and businesses via API, with a clear and competitive pricing structure. The cost is set at $5 per million input tokens and $25 per million output tokens. Tokens represent the fundamental units of information that an AI model processes, both when receiving instructions (input) and generating responses (output). For users requiring even faster processing, a "Fast mode" is available, operating at approximately 2.5 times the default speed, priced at $10 per million input tokens and $50 per million output tokens. This tiered pricing and readily accessible API ensure that businesses can leverage Opus 5’s capabilities based on their specific needs and budget.

Crucially, the release of Opus 5, available both via API and subscription for everyone, effectively resolves the uncertainty and anxiety that had surrounded Fable 5’s inconsistent public availability. It provides a stable, high-performance option for Anthropic’s paying users, solidifying the company’s product offering.

An interesting paradox emerges with Opus 5’s launch: despite Fable 5 being classified as part of the elite "Mythos-class" – a tier supposedly above Opus – Opus 5 consistently outperforms Fable 5 on nearly all metrics that matter to the everyday user. This raises questions about the practical utility of classification versus raw, demonstrable performance, suggesting that for many applications, the "Opus" line now represents the true frontier of accessible, high-performance AI.

The Broader AI Landscape and Competitive Pressures

The release of Claude Opus 5 unfolds within an intensely competitive and rapidly evolving global AI landscape. Anthropic, founded by former OpenAI researchers who prioritized AI safety and Constitutional AI principles, is a key contender in a market dominated by giants like OpenAI (backed by Microsoft) and Google. The direct comparisons in benchmarks against OpenAI’s GPT-5.6 Sol underscore the ongoing rivalry and the relentless pace of innovation driving these companies to push the boundaries of AI capabilities.

Beyond the established Western players, emerging forces are also making significant strides. Just a week prior to Anthropic’s announcement, Moonshot AI, a Beijing-based startup backed by Alibaba, unveiled Kimi K3. This model, boasting an impressive 2.8 trillion parameters, is notable for being an "open-weight" model, meaning its underlying code can be downloaded and run independently by anyone. This open-source approach contrasts with the proprietary models offered by Anthropic and OpenAI and carries significant implications for transparency, community-driven development, and broader accessibility of advanced AI. While independent benchmarks generally place Kimi K3 third overall, behind both Fable 5 and GPT-5.6 Sol, it has demonstrated superior performance in specific niche areas, highlighting the diverse strengths emerging from different AI development philosophies and geographical regions. The rise of models like Kimi K3 emphasizes the global nature of the AI race and the increasing importance of open-source contributions to the ecosystem.

Implications and Future Outlook

The launch of Claude Opus 5 carries profound implications for Anthropic, its user base, and the broader AI industry. For Anthropic, it represents a pivotal moment to regain trust and stabilize its product offering after the Fable 5 setback. By delivering a demonstrably superior and more reliable model at a competitive price point, Anthropic strengthens its position as a serious and responsible player in the generative AI market, reinforcing its commitment to high-performance AI grounded in safety.

For businesses and developers, Opus 5 offers access to a more capable, stable, and cost-effective tool for complex tasks ranging from advanced coding and scientific research to sophisticated business automation. The emphasis on reduced variance and consistent performance will be particularly attractive to enterprises looking to integrate AI reliably into critical operations. The model’s ability to handle multi-step analyses and execute full workflows autonomously signifies a step change in AI’s utility, moving beyond mere conversational agents to powerful operational partners.

For the AI industry as a whole, Opus 5’s introduction intensifies the ongoing competition, driving further innovation and pushing the envelope of what LLMs can achieve. It also underscores the continued tension between raw performance, accessibility, and the critical importance of AI safety and responsible deployment, especially in the wake of Fable 5’s regulatory challenges. The swift government intervention regarding the "jailbreak vulnerability" serves as a potent reminder that as AI models become more powerful, the need for robust safety mechanisms and effective regulatory frameworks will only grow.

In conclusion, Claude Opus 5 is more than just another model release; it is a strategic repositioning for Anthropic, a robust solution for developers, and a significant data point in the global AI race. By delivering on performance, reliability, and cost-effectiveness, Opus 5 is poised to become a foundational tool for a wide array of advanced applications, marking a new chapter in Anthropic’s journey to build safe and useful artificial general intelligence.

About the Author

About the Author

Easy WordPress Websites Builder: Versatile Demos for Blogs, News, eCommerce and More – One-Click Import, No Coding! 1000+ Ready-made Templates for Stunning Newspaper, Magazine, Blog, and Publishing Websites.

BlockSpare — News, Magazine and Blog Addons for (Gutenberg) Block Editor

Search the Archives

Access over the years of investigative journalism and breaking reports