Claude Opus 5 vs Claude Fable 5: Near-Frontier at Half the Price (2026)

Quick answer. Claude Opus 5 matches or beats Fable 5 on most benchmarks — Frontier-Bench (43.3% vs 33.7%), GDPval-AA v2 (1,861 vs 1,747), SWE-bench Pro within a point — at half the input price ($5 vs $10 per million tokens). Pick Fable 5 only for the longest-horizon autonomous agent work.

Anthropic shipped Claude Opus 5 on July 24, 2026 — its fourth model in under two months. The pitch is unusually blunt for a frontier lab: a model that "comes close to the frontier intelligence of Claude Fable 5 at half the price." That is a real claim you can check against numbers, and mostly it holds up.

This piece is the buyer's-eye comparison. Not a spec dump — a decision. When do you reach for the cheaper Opus 5, and when is the flagship Fable 5 still worth double the input cost? We lead with where Opus 5 wins outright, then stay honest about the one place Fable 5 still earns its premium.

What is Claude Opus 5, and how is it positioned against Fable 5?

Opus 5 is Anthropic's everyday workhorse for coding, agentic tasks, and knowledge work. It ships with a 1M-token context window (default and maximum — there's no smaller variant), extended thinking on by default, and a low/medium/high effort toggle that lets you trade latency and cost against reasoning depth per request. Standard output is 128K tokens, or up to 300K through the Message Batches API with a beta header.

Fable 5 remains the flagship. Anthropic still recommends it for the most advanced, long-horizon autonomous work — agents that run for days without a human in the loop. And Mythos 5 sits above both for frontier cybersecurity exploitation and the hardest biology research. So the family is a genuine ladder, not a rebrand: Opus 5 is the value tier that happens to touch the frontier on most tasks, not on all of them.

The headline is the price gap. Opus 5 keeps the same pricing as Opus 4.8 — $5 input / $25 output per million tokens — while roughly doubling Opus 4.8's benchmark scores. Fable 5 charges $10 per million input tokens, exactly twice Opus 5. That 2x delta is the whole decision, so let's put the benchmarks next to it.

How do Opus 5 and Fable 5 compare on benchmarks?

On the numbers Anthropic and independent trackers published, Opus 5 wins or ties on almost everything that matters for coding and knowledge work. The one exception is SWE-bench Pro, where it lands under a point behind the leaders.

BenchmarkOpus 5Fable 5GPT-5.6 SolWinner
Frontier-Bench v0.1 (agentic coding)43.3%33.7%34.4%Opus 5
GDPval-AA v2 (knowledge work, human-graded Elo)1,8611,7471,736Opus 5
ARC-AGI-3 (novel problem-solving)30.2%not listed7.8%Opus 5 (~3x next-best)
SWE-bench Pro79.2%80.0%Fable 5 / Mythos 5 (80.3%)

A few things stand out. On Frontier-Bench — agentic coding, the closest proxy to real engineering work — Opus 5 doesn't just edge Fable 5, it beats it by nearly ten points and more than doubles Opus 4.8 at a lower cost per task. On GDPval-AA v2, the human-graded knowledge-work Elo, Opus 5 leads the board outright. On ARC-AGI-3 it posts roughly triple the next-best model's score. Those are the everyday jobs most teams are actually buying a model for.

The gap closes on SWE-bench Pro, the hardest real-repo software task: Opus 5 scores 79.2%, third behind Mythos 5 (80.3%) and Fable 5 (80.0%) — but far ahead of Opus 4.8's 69.2%. In practice that's a rounding-error difference against a 2x price cut. Anthropic also reports Opus 5 lands within 0.5 points of Fable 5's peak on CursorBench 3.2 (max effort) at about half the cost per task, and beats Fable 5's peak on OSWorld 2.0 computer-use at roughly a third of the budget. On Zapier's AutomationBench it completed a full end-to-end churn-prevention workflow (100%) at about 1.5x the throughput of the next-closest model at matching cost.

If you're evaluating models for a coding agent stack specifically, our AI coding agents guide covers how these scores translate into agent reliability day to day.

How much cheaper is Opus 5 than Fable 5?

The pricing is the reason this comparison exists. Here's the full picture, including Opus 5's optional fast mode.

Model / modeInput ($/M tokens)Output ($/M tokens)Notes
Opus 5 (standard)$5$25Same as Opus 4.8; extended thinking on by default
Opus 5 (fast mode)$10$502x price, ~2.5x faster
Fable 5$102x Opus 5's input

So the core trade is simple: for the same input cost as Fable 5, you can run Opus 5 in fast mode — 2x price, ~2.5x the speed — and still come out ahead on Frontier-Bench and GDPval. Or you run standard Opus 5 at half Fable 5's input price and pocket the difference. Either way, the effort toggle (low/medium/high) is Anthropic's explicit lever to balance cost and capability: drop to low effort for routine, high-volume calls; push to high effort for the handful of genuinely hard prompts.

For high-volume workloads — a support-triage pipeline, a batch document job — halving the input rate while gaining benchmark points pays for itself fast.

Where does Fable 5 still beat Opus 5?

Honesty matters here, because the marketing line ("near-frontier at half the price") can read as "just buy the cheap one." That's not quite right. Fable 5 still wins in specific, high-stakes lanes:

  • Long-horizon autonomous agents. This is the explicit carve-out. For agents that run for days, planning and self-correcting across very long trajectories, Anthropic still recommends Fable 5. If your product is a multi-day autonomous worker rather than a request/response assistant, the flagship's extra headroom is worth the premium.
  • The hardest software tasks. Fable 5 (and Mythos 5) edge Opus 5 on SWE-bench Pro. It's under a point, but if your workload is dominated by the most brutal real-repo problems, that margin can compound.
  • Frontier cybersecurity and the hardest biology. These belong to Mythos 5, not Fable 5 — but the point stands that Opus 5 is not the top of the ladder for exploitation-grade security work. Opus 5 tracks close to Mythos 5 on vulnerability discovery (OSS-Fuzz) and can examine source code, but it won't scan compiled binaries and trails on exploitation.

For everything else — most coding, most agentic workflows, most knowledge work, computer use, scientific research (Opus 5 is Anthropic's most capable generally available model for research, improving on Opus 4.8 across all evals with organic chemistry up 10 points) — Opus 5 is the rational default.

Do data retention and refusal rates differ between the two?

Yes, and these are underrated differentiators for anyone shipping to production or handling sensitive data.

Data retention. Opus 5 is not subject to the 30-day data-retention policy that applies to Fable 5. For teams with compliance requirements or client contracts that restrict where prompt data lingers, that's a concrete reason to prefer Opus 5 independent of price or benchmarks.

Refusal rate. Opus 5's safety classifiers are expected to trigger about 85% less often than Fable 5's — meaning far fewer spurious refusals on legitimate work. Anyone who's had a model refuse a perfectly reasonable security-research or content-moderation prompt knows how much friction that removes. Anthropic's own automated behavioral audit also scores Opus 5 as its most aligned model to date (misaligned-behavior score 2.3, the lowest of recent Claude models, with the lowest deceptive-behavior reading), so the reduced refusals don't come from loosened safety — they come from better calibration.

There's also a new beta "Automatic Fallbacks" safeguard on the Platform: when a safety classifier trips, instead of returning an error, the API routes to a less-powerful model and hands back a usable response. Fewer hard failures in your pipeline.

Which model should you actually pick?

A short decision rule:

  • Default to Opus 5 for coding assistants, request/response agents, knowledge work, computer use, scientific research, and any high-volume workload where the halved input price compounds. It self-verifies and recovers from its own errors without user intervention, so you spend less time babysitting it.
  • Reach for Fable 5 when you're building genuinely long-horizon autonomous agents that run for days, or when your workload lives at the very top of SWE-bench-Pro difficulty and that sub-point margin matters.
  • Reach for Mythos 5 only for frontier cybersecurity exploitation or the hardest biology research.
  • Use Opus 5 fast mode when you'd otherwise pay Fable 5's rate anyway and you want speed — you get ~2.5x throughput at the same input price, still ahead on the headline benchmarks.

If you're also weighing Anthropic against OpenAI's latest, our GPT-5.6 vs Claude Fable 5 comparison sits alongside this one — and note that Opus 5 beats GPT-5.6 Sol on Frontier-Bench (43.3% vs 34.4%), GDPval (1,861 vs 1,736), and ARC-AGI-3 (30.2% vs 7.8%) at a lower price point. For a cheaper still tier below Opus, Claude Sonnet 5 covers the lightweight end of the family.

Building on top of these models and need engineers who already know the ecosystem? Codersera places vetted remote developers who ship production AI and agent systems — you can extend your team with Codersera and skip the months of hiring risk. Faster technical fit, lower hiring risk, remote-ready from day one.

FAQ

Is Claude Opus 5 actually as good as Fable 5?

On most benchmarks, yes — Opus 5 beats Fable 5 on Frontier-Bench (43.3% vs 33.7%) and GDPval-AA v2 (1,861 vs 1,747), and lands within a point on SWE-bench Pro (79.2% vs 80.0%). Fable 5 pulls ahead specifically on long-horizon autonomous agent work and the very hardest software tasks.

How much cheaper is Opus 5 than Fable 5?

Opus 5 costs $5 per million input tokens versus Fable 5's $10 — exactly half. Output is $25 per million. Opus 5 also offers a fast mode at $10 input / $50 output (2x price, ~2.5x faster), which matches Fable 5's input rate while running significantly quicker.

When should I still use Fable 5 instead of Opus 5?

Choose Fable 5 for the most advanced long-horizon autonomous agents — those running for days without human intervention — where Anthropic still recommends it. It also edges Opus 5 on the hardest SWE-bench Pro problems. For nearly everything else, Opus 5 is the better value.

Does Opus 5 have the same data-retention policy as Fable 5?

No. Opus 5 is not subject to the 30-day data-retention policy that applies to Fable 5, and its safety classifiers trigger roughly 85% less often, meaning fewer spurious refusals. Both are meaningful reasons to prefer Opus 5 for sensitive or production workloads.

What is Opus 5's effort toggle?

Opus 5 exposes a low/medium/high effort parameter per request that controls how much reasoning compute the model spends. Low is faster and cheaper for routine tasks; high spends longer reasoning on hard problems. It's Anthropic's lever for balancing cost against capability without switching models.