The AI Teardown· August 5, 2026
Claude Fable 5’s Fallback: Fable Prices, Opus Answers
The most expensive model on the market does not always answer when called. Claude Fable 5 costs $10 in and $50 out per million tokens, and Opus 4.8, the model it hands flagged work to, costs half that.
Anthropic’s line is that fallback occurs in fewer than 5% of sessions. Artificial Analysis recorded Fable falling back on roughly 8% of its Intelligence Index tasks, rising to 9% on graduate-level science.
Both figures are averages, and the averages are the least informative numbers in this story. The rate you face follows your workload, and for some workloads it is not a rate at all. It is a wall.
The rate you face is not 8%
The fallback follows the safety classifiers, which means it follows what you do for a living. Artificial Analysis measured 2% on real-world work tasks against 9% on hard science, and Vals AI measured near-total refusal on biology and cybersecurity questions.
Counting those refusals as failures collapses Fable’s GPQA Diamond score from 93.18%, second place, to 55.56%, ninety-fourth. Same model, same test: second-best in the world or barely mid-table, depending entirely on whether the questions it declined to answer count against it.
One stock-harness run made it concrete. Fable refused 26 tasks that Opus 4.8 simply completed, including four defensive security reviews, five routine bioinformatics jobs, and one literature review on AI-assisted drug discovery. A model Anthropic markets for finding vulnerabilities in critical software declined to audit a Flask app for the developer who owns it, as “violative cyber content”.
So the honest per-operator maths: ordinary coding traffic rarely sees the fallback. Workloads touching security review or the life sciences, the exact domains a frontier premium is supposed to buy, run toward total.
The bill arrives either way
The premium is charged whether or not the flagship shows up. Running Humanity’s Last Exam on Fable cost Artificial Analysis roughly $2.2k, the highest bill of any model it has ever evaluated, fallback costs included.
Subscription users hit the same wall in their session windows: one Fable prompt eating 20% of a five-hour Max session, one code review of a 64MB iOS app consuming a full daily budget before it finished. One benchmark run priced the gap exactly, at $8.39 on GPT-5.6 Sol against $21.63 on Fable for the same work.
Session windows are the real pricing page. The per-token rate card is decoration.
The paying cohort has noticed, and its mood is not loyalty. The r/ClaudeAI thread warning Anthropic to keep Fable on subscription, “otherwise, we’ll downgrade and bail in mass”, sits at 80 votes, every one of them from a card on file.
The sellers already route around the flagship
The labs’ own defaults answer the question their charts avoid. Claude Opus 5 shipped on 24 July at $5 in and $25 out, within a point of Fable on repository work at half the output price, and Anthropic made it Claude Code’s default the same day. OpenAI had already made Sol, not its priciest configuration, the Codex default on 9 July.
The vendors price their own flagships as special-occasion models, and the defaults they chose are the honest chart.
The check most operators never run
Unsurprisingly, the answer has been sitting in the documentation the whole time. Anthropic’s own system card says the apps display a notice when a query is rerouted, and the API records every substitution in the response object, with a session event firing each time.
Which model actually served each request is already written down, per request, in machine-readable form. Reading it is one line of code.
An operator who reads it knows, workload by workload, whether the premium is arriving or whether a tenth of the bill is buying the understudy. The routing decision then writes itself: the flagship for work that clears its classifiers and earns the two-to-one price, the half-price near-peer for everything else.
By the numbers
- Under 5%: Anthropic’s stated fallback rate, per session
- 8 to 9%: the measured rate on Intelligence Index tasks and hard science
- 93.18% to 55.56%: Fable’s GPQA Diamond score with refusals counted against it
- 26: tasks Fable refused that Opus 4.8 simply completed
- $21.63 vs $8.39: Fable against Sol for the same benchmark work
- 80 votes: the downgrade ultimatum, all paying subscribers
The close
The flagship reserves the right to send its understudy, and the receipt says which one showed up. Operators reading their response objects this month will route the premium to the work that earns it and pocket the difference. Everyone else keeps paying Fable prices for Opus answers, and Anthropic keeps the change.