Ranked #3 Everyday Ecosystem — The Leading AI Assistants
Anthropic

Claude Fable 5.1

Same weights as the invite-only Mythos 5.1, but with the safety layer most people actually get. It's the model Anthropic tells you to use when Opus 5 at high effort still loses your private eval — long-horizon coding, multi-hour research, docs/sheets/slides that have to stay coherent.

ReasoningLong-ContextAgentic
9.8out of 10
Official Website
Best for

Same weights as the invite-only Mythos 5.1, but with the safety layer most people actually get. It's the model Anthropic tells you to use when Opus 5 at high effort still loses your private eval — long-horizon coding, multi-hour research, docs/sheets/slides that have to stay coherent.

Why It Wins

Artificial Analysis Intelligence Index 66 at max (highest measured at launch). GDPval-AA v2 1853. Cache reads cut 75% to $0.25/MTok — the real price change. 1M context, 128k out, adaptive thinking always on.

Watch out

Sticker price is still $10/$50 — 2× Opus 5. Max costs $3.76/task vs Opus $2.34 because it writes ~1.7× more tokens. Max effort can over-edit. Covered Model (30-day retention). No native image gen.

01

What It Actually Is

If Opus 5 is the reliable workhorse you trust for 90% of your daily knowledge work, Fable 5.1 is the specialist you hire for the remaining 10% that actually keeps you up at night. Anthropic isn’t introducing a brand-new architecture here; Fable 5.1 is a point release that tightens the classifiers on the original Fable 5 and, most importantly, slashes the cost of caching context. It shares the same weights as the highly restricted Mythos 5.1, but comes with the safety layer that makes it generally available to the public.

The capability jump is real, but it’s measured in stamina rather than raw peaks. Fable 5.1 took the #1 spot on the Artificial Analysis Intelligence Index at launch with a score of 66 at max effort, beating out Opus 5 (63) and GPT-5.6 Sol (61). It also posted a dominant 1853 on GDPval-AA v2. But the most significant change isn’t in the neural network—it’s on the pricing page. While the base token costs remain a premium $10/$50, cache read prices have been cut by 75% down to $0.25 per million tokens. For workflows that depend on a hot 1M context window—like querying massive codebases or synthesizing hundreds of documents—this fundamentally changes the economics of using a frontier reasoning model.

However, using Fable 5.1 is not a free lunch. The model features always-on adaptive thinking, which means it will chew on a problem until it’s satisfied. The result is intense token hunger; Fable 5.1 writes roughly 1.7× more tokens per task than its predecessor. On Artificial Analysis, max effort costs $3.76 per task compared to Opus 5’s $2.34. And ironically, more effort doesn’t always equal better results. The system card reveals that max effort can lead to “over-editing”—adding unnecessary documentation, drive-by file tweaks, or complex CI jobs where a simple fix would do.

There are also the operational realities of a Covered Model. Fable 5.1 retains your data for 30 days unless you have Enterprise Frontier Safeguards (EFS) or Zero Data Retention (ZDR) agreements in place. Furthermore, the safety classifiers remain aggressive. While it can identify vulnerabilities in your code, it will refuse to write exploits, and complex biological research queries will silently trigger a fallback to the Opus model. And if you frequently hit the API with a cold 1M context, those $10/MTok cold reads will punish your usage caps quickly.

Anthropic’s own routing advice is the best summary of Fable 5.1’s place in the ecosystem: start with Opus 5. But when Opus fails your private evals—when you’re dealing with multi-hour research, long-horizon coding tasks, or complex documents that must maintain strict coherence—Fable 5.1 is the most capable fallback in the world. It is the model you pay extra for when getting it right the first time is the only thing that matters.

02

Strengths and honest limitations

Key Strengths

  • AA Intelligence Index 66: At launch, Fable 5.1 took the #1 spot on Artificial Analysis Intelligence Index with a score of 66 at max effort, outperforming Opus 5 and GPT-5.6 Sol.
  • Massive Cache Read Discount: While the sticker price remains high, cache reads are cut by 75% to $0.25/MTok. For workflows relying on a hot 1M context, this is a game-changing price reduction.
  • Adaptive Thinking Always On: The model features always-on adaptive thinking, giving it the time it needs to reason through complex, multi-step problems that trip up single-pass models.
  • Long-Horizon Mastery: With 1M context and 128k output limits, it handles multi-hour research tasks and massive document processing while maintaining coherence better than any competitor.
  • Composite Intelligence Leader: Despite Astra’s lead in specific categories like computer-use and math, Fable 5.1 retains the crown for composite intelligence and HLE (Human Level Evaluation).

Honest Limitations

  • Token Hunger & Higher Costs: Fable 5.1 writes ~1.7× more tokens than Fable 5. On Artificial Analysis, max effort costs $3.76/task compared to Opus 5’s $2.34. The cache cut helps, but it’s not a free lunch.
  • Over-Editing at Max Effort: More thinking isn’t always better. The system card reveals that ‘medium’ effort often beats ‘max’ in software engineering because max adds drive-by file edits, extra docs, and unnecessary CI jobs.
  • Covered Model Restrictions: Fable 5.1 operates as a Covered Model with 30-day retention (unless EFS/ZDR). The unrestricted Mythos 5.1 is locked behind trusted access only.
  • Safety Classifiers Still Eat Dual-Use: The model can find vulnerabilities but will not write exploits. If you push into bio research, the safety system will automatically route you to Opus.
  • Cold Cache Penalty: Claude usage caps punish users who frequently load a cold 1M context. The full $10/MTok applies to cold reads, eating into usage limits quickly.
03

Benchmark Snapshot

Artificial Analysis Intelligence Index — 66 (Max)

Highest measured at launch, ahead of Opus 5 (63) and GPT-5.6 Sol (61).

GDPval-AA v2 — 1853

Outscores Fable 5 (1723) and Opus 5 (1824).

HLE tools — 65.0%

Leads the pack in Human Level Evaluation with tools, beating Opus 5 (63.6%) and Astra (57.2%).

Signal65 PINNACLE — 7,898

276/280 agent jobs completed; $2.46 per correct task.

04

The Verdict

Claude Fable 5.1 isn’t a new family — it’s a precision update on Fable 5 that slashes cache read costs and tightens the classifiers. While Anthropic’s own rule is to ‘start with Opus 5’, Fable 5.1 is the heavyweight specialist you tag in when Opus fails. The Artificial Analysis Intelligence Index score of 66 proves its raw capability, and the 75% cache read discount makes long-context workflows viable. But the high sticker price, token hunger, and aggressive safety routing mean it’s not for everyday chatter. Use it when the job demands multi-hour research, complex coding, or coherent generation across a 1M token context, and you’re willing to pay the premium for the smartest agent in the room.