AIToday
Large Language ModelsAI Business & IndustryTHE DECODERPublished: Jul 25, 2026, 19:01 JST3 min read

Claude Opus 5 outperforms Fable 5, costs less on most benchmarks

Claude Opus 5 outperforms Fable 5, costs less on most benchmarks

3 Key Points

  1. What happened

    Anthropic released Claude Opus 5, which scored 61 on the Artificial Analysis Intelligence Index—ahead of Fable 5 (60)—and matched or beat Fable 5 on most benchmarks including coding, scientific reasoning, and knowledge work tasks. Token pricing remains at $5 per million input tokens and $25 per million output tokens.

  2. Why it matters

    Opus 5 costs less than Fable 5 on average Intelligence Index tasks ($2.03 vs. $2.75) and substantially less on knowledge work—$10.41 at the "high" reasoning tier versus $22.30 for Fable 5—while delivering superior analytical quality (Analytical Quality Elo of 2016 vs. Fable 5's ~1700). This suggests frontier models are becoming more cost-competitive rather than pulling away from each other.

  3. What to watch

    The "high" reasoning tier emerges as the best value for Opus 5 on coding tasks, beating the costlier "max" tier due to time constraints limiting attempts. At "max," Opus 5 takes over 36 minutes per task on knowledge work, about 50% longer than Opus 4.8.

Ask the AI about this article →

Summaries like this, in your inbox every morning.

Context & Analysis

Claude Opus 5 arrives at a point of convergence among frontier AI models. Both Artificial Analysis and Epoch AI confirm that the performance gap between leading models has tightened considerably. Artificial Analysis tested Opus 5 at 61 points on its Intelligence Index, just one point ahead of Fable 5 and only two ahead of GPT-5.6 Sol. Epoch AI's independent assessment reinforces this picture: Opus 5 scores 159 on the overall Epoch Capability Index versus Fable 5 at 161, a margin too small to claim superiority. The article explicitly notes that "no single model can pull away or claim a clear advantage," suggesting that the era of incremental model releases each delivering decisive improvements may be ending.

What sets Opus 5 apart is not breakthrough capability but cost-effectiveness and specialization. On knowledge work—tasks like writing reports, building presentations, and analyzing data—Opus 5 achieves dramatic cost advantages. At the "high" reasoning tier it costs $10.41 per task versus $22.30 for Fable 5 while beating Fable 5 in Elo ranking. Its analytical quality is particularly strong, reaching an Analytical Quality Elo of 2016 at "max," nearly 300 points ahead of Fable 5. This pattern suggests that rather than all-purpose superiority, Opus 5 excels in specific domains and pricing configurations. The body's data on reasoning tiers reveals an interesting tension: higher reasoning tiers produce more complex (and error-prone) solutions, while the default "high" tier strikes a reliability-to-cost balance. This insight aligns with Anthropic's own API design, which sets "high" as the default—a practical acknowledgment that more computation does not always yield better results.

FAQ
How much does Claude Opus 5 cost compared to other models?
Token pricing is $5 per million input tokens and $25 per million output tokens. On knowledge work tasks at the "high" reasoning tier, it costs $10.41 per task versus $22.30 for Fable 5. On average Intelligence Index tasks, it costs $2.03 versus $2.75 for Fable 5.
Which reasoning tier should I use for best value on coding tasks?
The "high" reasoning tier delivers the best results for coding tasks. Vals.ai found that the highest tiers ("xhigh" and "max") produce more complex solutions that contain errors more often, while "high" produces simpler solutions that meet requirements more reliably.
Where does Opus 5 fall short compared to rivals?
Factual accuracy remains a weak spot. On AA-Omniscience, Opus 5 trails Fable 5, and its hallucination rate is 50 percent. On presentation quality for knowledge work, it scores a Presentation Elo of 1628, about 40 points behind GPT-5.6 Sol at "max" (1666).

Get the latest Large Language Models news every morning

For example, today's edition would include:

  • ChatGPT Work with GPT-6 Astra builds 5K running loop in 27 minutesSimon Willison's Weblog · 2h ago
  • OpenAI agents linked to RubyGems attack in MayThe Verge AI · 2h ago
  • Tesla Reportedly Pushes Staff Toward Grok 4.5 as AI Spending Cap Takes EffectTop Companies AI · 6h ago

AI-summarized, only the topics you pick — one digest a day via Email, Slack, or Discord.

Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →

Ask AI

Ask AI anything about this article. Q&As are published on this page for other readers too.

Related Articles

Next articleCapital One bets $125M on MLB to win credit card loyalty wars