Anthropic's Claude Opus 5 has secured the top position on MLQ.ai's AI benchmark index while carrying a price tag that sits at roughly half what users pay for OpenAI's Fable 5. The results, published this week, add weight to the argument that frontier AI performance and accessible pricing no longer have to be in tension.

Benchmark Performance

MLQ.ai's index evaluates models across a range of tasks including reasoning, coding, instruction following, and agentic workflows. Claude Opus 5 posted leading scores across several of those categories. In agentic search specifically, Opus 5 has shown a consistent edge over Fable 5, a pattern the MLQ.ai data appears to reinforce. The margin was not trivial; the model outperformed competitors by a measurable degree on multi-step reasoning tasks that have historically been a weak point for large language models.

Key Facts

  • Claude Opus 5 ranks first on MLQ.ai's AI benchmark index
  • Priced at approximately half the per-token cost of Fable 5
  • Leads on agentic and multi-step reasoning tasks
  • Benchmarks cover coding, reasoning, and instruction following
  • Results published by MLQ.ai following the model's launch

The pricing gap is significant for enterprise buyers who run models at scale. At half the cost of Fable 5, organizations processing millions of tokens per day could see their inference bills drop sharply without sacrificing output quality. That calculus has already prompted conversations among AI procurement teams about whether the default choice for high-stakes workloads should shift. Anthropic built a cost-capability toggle into Opus 5, giving operators a direct mechanism to balance quality against spend depending on the task at hand.

"Topping a major benchmark index while pricing below the nearest competitor is an unusual combination. Historically, the most capable models have also been the most expensive."MLQ.ai analysis
Claude AI Handboek by Leon Tindemans
Get the Claude AI Handboek
458 pages on getting more out of Claude, by AI expert Leon Tindemans. A printed book, written in Dutch, shipped worldwide with track and trace.
View the book →

Where Opus 5 Fits in the Market

The launch arrives at a moment when the frontier model tier is becoming genuinely crowded. Alibaba's Qwen3.8 recently claimed a second-place position behind Fable 5 on separate evaluations, suggesting that strong challengers are emerging from multiple directions. Opus 5's arrival reshuffles that picture considerably. For Anthropic, the benchmark result is a public validation of the engineering decisions baked into Opus 5's training and architecture.

It is worth keeping in mind that no single benchmark tells the full story of a model's usefulness. MLQ.ai's index carries credibility in the developer community, but real-world performance varies depending on deployment context, prompt design, and the specific domain a team is working in. Enterprises evaluating Opus 5 are likely to run their own internal evals before committing at scale. Still, topping an independent index while cutting the price of the category leader is a strong opening position.

For developers and organizations tracking where the technology is heading, the Opus 5 data points to a broader trend: the gap between cost and capability is compressing faster than most analysts predicted at the start of the year. Whether that trajectory holds through the next generation of releases remains to be seen, but for now Claude Opus 5 sits at the front of a competitive field. You can follow the latest Claude AI news for ongoing coverage as independent evaluations continue to roll in.

Further reading: Learn more about Claude's model family, read our background on Anthropic, or browse the latest Claude AI news.