LatentEval

For everyday use

Do you need Fable 5? A task-by-task verdict, every number sourced

The free window closed on July 19, 2026. Fable 5 is now included on Max and premium Team seats, and runs on usage credits for Pro. Which tasks justify it, and which belong on Opus 4.8.

For everyday use

In brief

5 POINTS
  • Fable 5 earns its price only on long, multi-step agent runs; everyday chat and short coding belong on Opus 4.8.
  • On benign coding, Fable 5 declined 20 of 28 tasks and Claude silently served them from Opus 4.8, so you paid double for the detour.
  • Our own test of short, checkable tasks could not measure Fable 5 beating Opus 4.8, so the premium is insurance bought sight unseen.
  • On Max and premium Team seats, Fable 5 covers at most half your weekly usage before drawing usage credits.
  • The free window closed on July 19, 2026, so on Pro and standard Team seats every Fable 5 call now draws usage credits.

Open Claude and Fable 5 sits at the top of the model list. Leave it selected and it will draft your emails, tidy your paragraphs, and patch your one-line bugs. The picker shows you a model; it never shows you the price. That same habit bills at twice Opus 4.8’s rate2, on work Opus was already doing faster and for half the money. You would never notice. The bill would.

Route by task: switch Fable 5 on for long, multi-step research and agent runs, and turn it off for everyday chat, drafting, and short coding, where Anthropic itself points you to Opus 4.82 at half the price. Backwards, and it costs you twice over: double the rate, on work Opus already nails, for a gain you never feel.

Max plans and premium Team seats, since July 20, 2026: Fable 5 is included for up to half your weekly usage limits, with no end date1. Pro plans and standard Team seats: it runs on usage credits (the metered balance your plan bills against) or the standard metered rate, at twice what Opus 4.8 costs2.

What is Claude Fable 5?

Claude Fable 5 is Anthropic’s top-tier Claude model, released alongside Claude Mythos 5 and pointed at the most demanding reasoning and long-horizon agent work.3 It sits one rung above Opus 4.8 on both capability and cost, listing at $10 per million input tokens and $50 per million output, twice Opus 4.8’s $5 and $25.3 Anthropic’s own steer is to keep everyday work on Opus 4.8 and reach for Fable 5 only when a job needs the highest available capability.2 The rest of this page turns that steer into a task-by-task verdict, every price and date sourced.

How the free window worked, and who it left out

Anthropic pulled Fable 5 on June 12 under a US export-control order, then switched it back on worldwide on July 1. The version that returned came with a fuse. On Pro, Max, Team, and select Enterprise plans, it ran free for at most half your weekly usage and, as announced at the time, only through July 74; crossing either line put you on usage credits. That July 7 date was extended twice, as the note above sets out.

The word “included” did not mean the same thing on every plan. Standard Enterprise got no included allowance4: there, Fable 5 ran on usage credits from the first message. Everyone else (Pro, Max, Team, and select Enterprise) held the capped free window, which in the end closed on July 19.

Is Fable 5 actually better for your work?

Anthropic calls Fable 5 its most capable widely released model, built for the most demanding reasoning and long-horizon agentic work: jobs that chain many steps toward one result.3 Whether the bill is worth paying comes down to a single question: is your work demanding in that specific way?

Two panels: in chained work an early error is inherited and grows down the steps; standalone tasks keep the same error in its tile, never spreading.
Chained work compounds an early error; standalone tasks contain it. The verdict turns on one question: does your work chain? Explanatory diagram, no data

Fable 5 is a senior specialist you fly in for the day. On a weeks-long project, that judgment pays for itself many times over. On a fifteen-minute question, you buy the whole day and get the answer the colleague at your desk already had.

Picture Maya, nine hours into a literature review across dozens of papers, feeding Claude each new source to reconcile against everything it has already summarized. The latest source never gets read fresh; the model reasons over the running summary that every earlier source built.

Each step builds on the one before it. A sharper reconciliation at paper six becomes the ground paper seven stands on. A small error there is inherited as settled fact by every source that follows. That compounding is where a stronger model earns its price. A single email gives it nothing to build on.

Marcus, opening Claude the same morning, is fixing the tone of a cover letter and a function that throws every time the input list arrives empty. Two quick, self-contained jobs, with nothing downstream waiting on either. For exactly that work, Opus 4.8 is the model Anthropic routes you to2, and it answers faster at half the metered price. Maya switches it on; Marcus leaves it off.

What our own test found

We put that to our own test. On short tasks with a checkable right answer (structured extraction, tool calls, code you can unit-test), we could not measure Fable beating Opus at all: the two land inside the same margin at both effort settings we tried.5 The work Fable is built to win is the other kind, the long, open-ended, many-step runs, and that is the work this sort of test cannot reach. So the edge you pay for is real in theory and, on everyday checkable work, unmeasured in practice. It is insurance you buy sight unseen.

Anthropic draws the line the same way: start with Opus 4.8 for most work, and reach for Fable 5 only for workloads that need the highest available capability.2 If each step of your task feeds the next, Fable 5 can move the outcome; if your prompts stand alone, it cannot, and you would pay double for nothing.

With the free window gone, the bill doubles for the same prompt

All the date ever did was remove the cover. Underneath, the price was always double. Claude bills per token, the chunks of text a model reads and writes, and Fable 5 lists at $10 per million in and $50 out3, against Opus 4.8’s $5 and $25. For every dollar Opus charges, Fable 5 charges two.

Claude Fable 5Claude Opus 4.8
API price per 1M tokens$10 in / $50 out$5 in / $25 out
Latencyslower, you wait while it spins upmoderate
Best fitlong, multi-step reasoning and agent runseveryday chat, drafting, short coding

Prices and the latency row come from Anthropic’s model overview2; the best-fit split from that page’s routing guidance and Fable 5’s own description3.

On an everyday request, Fable 5 makes you wait longer and costs more for the answer Opus already hands you. That is waste: the same answer for a higher bill. But does Fable 5 ever make everyday work actively worse?

On everyday coding, most declines route straight to Opus anyway

On one slice of quick coding, yes. Fable 5 ships a stricter safety filter, one that blocks more requests before answering, and Anthropic is candid about the trade: it comes at the cost of flagging benign requests more often during routine coding and debugging tasks4. Flagged inside Claude’s own apps, you still get an answer: you see a notice, and the request is handed to Opus 4.84.

Pictograph of 28 benign coding tasks: Fable 5 answered 8 itself and declined 20, which fallback silently served from cheaper Opus 4.8. You pay twice.
On everyday coding, Fable 5 hands 20 of 28 tasks back to Opus 4.8. Left to itself, Fable 5 declined 75 to 86% of 28 benign, checkable coding tasks (more thinking did not help). Inside Claude, fallback quietly served those declines from Opus 4.8. Source: LatentEval routing-eval (Anthropic API, 2026-07-02). n=28 short-horizon checkable coding tasks, low effort; fallback-on arm.

How often is “more”? In the same test, on benign, checkable coding tasks (parsers, state machines, string transforms, small simulations), Fable declined most of them. With fallback on, 20 of 28 rerouted to Opus; left to itself, Fable declined 21 of 28 at low effort and 24 of 28 at maximum, roughly 75 to 86% either way. More thinking held the rate right where it was, at both settings. Of the 45 declines, 44 were the filter reading ordinary code as a cyber risk. The answer a user actually received landed on about one try in six, 5 of 28 at low effort and 4 of 28 at maximum, against roughly nine in ten for Opus (25 to 27 of 28).

Set that beside the price. On a slice of ordinary coding help, the pricier model declines, then reroutes you to the cheaper one, and you read the Opus answer regardless. You land in the same place. You just paid Fable 5’s premium for the detour to an Opus answer you could have picked yourself.

Route the work, then check your plan

  • Switch it on for long, multi-step reasoning or agent tasks. That is where a stronger model moves the outcome.
  • Leave it off for short chats, quick drafts, and one-off code fixes. Anthropic routes that work to Opus 4.82, at half the API price and lower latency. The same capability-against-price call runs one tier down, where the choice between Opus and Sonnet turns on how much reasoning the job actually needs.
  • Check your plan. Max and premium Team seats include Fable 5 for up to half your weekly usage; Pro and standard Team seats run it on usage credits from the first message1.
  • Decide what it is worth. If you do keep Fable 5 on, settle whether the work is worth usage credits or the metered $10 in, $50 out3, because on Pro and standard Team seats paying is now the default.

The free window has closed. The routing choice underneath it has not changed: every time you open Claude, you are still picking a model. “Most capable” measures peak benchmark scores. Whether a model gets your specific job right, run after run, is a separate measurement, and our research desk covers that gap in how to measure agent reliability past a single pass rate. Pick the model your task needs, no matter what sits at the top of the leaderboard.

Footnotes

  1. Anthropic, Claude Fable 5 on your plan (Claude Help Center). The July 19, 2026 11:59:59 PM PT end of the included promotion, the July 20 permanent split, the 50% weekly cap on Max and premium seats, the usage-credit terms on Pro and standard seats, and the one-time credit for eligible seats: https://support.claude.com/en/articles/15424964-claude-fable-5-on-your-plan 2 3 4 5

  2. Anthropic, Claude models overview, https://platform.claude.com/docs/en/about-claude/models/overview 2 3 4 5 6 7 8 9

  3. Anthropic, Introducing Claude Fable 5 and Claude Mythos 5, https://platform.claude.com/docs/en/about-claude/models/introducing-claude-fable-5 2 3 4 5 6

  4. Anthropic, Redeploying Claude Fable 5, https://www.anthropic.com/news/redeploying-fable-5 2 3 4

  5. Our own routing-eval harness: deterministic exact-match / unit-test scoring (no LLM judge), Anthropic API, pricing re-verified live 2026-07-02; an effort sweep of 14 pruned tasks x 2 runs = 28 trials per arm (prior round n=40), 95% Wilson intervals, a pre-registered stop rule. It measures short-horizon checkable correctness only, not long-horizon agentic or open-ended quality. Full method, per-arm counts, confidence intervals, and the results CSV live in the canonical write-up: the routing eval. Method and definitions: agent reliability testing.