Model comparison
Claude Fable 5.1 vs Claude Opus 5.5: when to pay 2.5x
Claude Fable 5.1 lists at 2.5x the rates of Claude Opus 5.5. What that premium buys, why cheap cache hits barely narrow it, and when Anthropic suggests it.
· Prices as of September 26, 2026
Claude Fable 5.1
Anthropic · Released September 1, 2026
Anthropic's most capable generally available model, aimed at demanding reasoning and long-horizon agentic coding. Anthropic suggests it when Opus-tier results fall short.
Claude Fable 5.1 facts and comparisonsClaude Opus 5.5
Anthropic · Released September 22, 2026
Anthropic's recommended starting model for most work, built for long-running agentic coding. Anthropic says it matches Claude Fable 5.1 on most work at a much lower price.
Claude Opus 5.5 facts and comparisons
The short answer
Claude Opus 5.5 is Anthropic's recommended starting model and the default in Claude Code, and Anthropic says it performs at the level of Claude Fable 5.1 on most work. Fable 5.1 lists at 2.5x the Opus 5.5 rates, and the example agentic coding session costs $10.50 on it against $4.40, a 2.4x gap. Anthropic suggests Fable 5.1 for the cases where Opus-tier results fall short.
Choose Claude Fable 5.1 if
- Opus-tier results fall short on a task, which is the point where Anthropic suggests moving up to Fable 5.1.
- Your work is demanding reasoning or long-horizon agentic coding, the use Anthropic aims its most capable generally available model at.
- You want a higher starting effort: Fable 5.1 starts at high, where Opus 5.5 starts at medium.
Choose Claude Opus 5.5 if
- You want the model Claude Code picks by default on Pro, Max, Team, and Enterprise plans and with an Anthropic API key.
- You run long, sprawling jobs such as codebase-wide migrations and audits, which Anthropic lists as an Opus 5.5 strength, at $4 input and $20 output per million tokens.
- Your organization requires zero data retention, which rules out Fable 5.1 and its 30-day retention requirement.
- You send many uncached requests or long outputs, where the full 2.5x rate gap applies to every token.
Side by side
Specs and prices
| Fact | Claude Fable 5.1 | Claude Opus 5.5 |
|---|---|---|
| Maker | Anthropic | Anthropic |
| API model id | claude-fable-5-1 | claude-opus-5-5 |
| Released | September 1, 2026 | September 22, 2026 |
| Status | Current | Current |
| Context window | 1M tokens | 1M tokens |
| Max output | 128K tokens | 128K tokens |
| Open weights | No | No |
| Input, per 1M tokens | $10 | $4 |
| Cache hit, per 1M | $0.25 | $0.20 |
| Cache write, per 1M | $12.50 (5-minute), $20 (1-hour) | $5 (5-minute), $8 (1-hour) |
| Output, per 1M tokens | $50 | $20 |
| Runs in | Claude Code, Cursor, OpenCode, OpenRouter, and GitHub Copilot | Claude Code, Cursor, OpenCode, OpenRouter, and GitHub Copilot |
Standard API rates in US dollars, as published by each maker on September 26, 2026. Batch and priority tiers, taxes, and subscription plans are not included. Claude Fable 5.1: The full 1M context window is billed at standard rates. Claude Opus 5.5: The full 1M context window is billed at standard rates. Fast mode, a research preview on the Claude API, costs $8 input and $40 output per million tokens.
Cost
What the same work costs
The same token counts, priced at each model’s published rates. Cached tokens are billed at each maker’s cache prices, so the session shows what caching is worth on each model.
Real sessions differ: the two models count the same code as different numbers of tokens, and reasoning settings change how much each one writes. Your own history is the real test.
| Workload | Claude Fable 5.1 | Claude Opus 5.5 |
|---|---|---|
| Agentic coding session, 100K input, 400K written to cache, 2M read from cache, 50K output | $10.50 | $4.40 |
| Large one-off review, 150K input with no cache hits, 10K output | $2.00 | $0.80 |
| Output-heavy generation, 30K input, 80K output | $4.30 | $1.72 |
| A month of sessions, 110 sessions: 5 a day, 22 working days | $1,155.00 | $484.00 |
| Where the session’s cost goes | ||
| Cache writes | $6.50 | $2.60 |
| Cache reads | $0.50 | $0.40 |
| Uncached input | $1.00 | $0.40 |
| Output | $2.50 | $1.00 |
- caching saves on the session with Claude Fable 5.1 (62%)
- $17.00
- caching saves on the session with Claude Opus 5.5 (60%)
- $6.60
How much more does Claude Fable 5.1 cost than Opus 5.5?
Claude Fable 5.1 charges $10 per million input tokens and $50 per million output tokens. Claude Opus 5.5 asks $4 and $20 for the same tokens. Cache writes keep the same 2.5x ratio: $12.50 and $20 for 5-minute and 1-hour writes on Fable 5.1, against $5 and $8 on Opus 5.5. Any work that skips the cache pays the full multiple, so the large one-off review costs $2.00 against $0.80 and the output-heavy generation $4.30 against $1.72.
Cache hits are where the two models come closest. Anthropic bills a hit at 0.025x the input price on Fable 5.1 and 0.05x on Opus 5.5, so a cached million tokens costs $0.25 on one and $0.20 on the other, a 1.3x gap. On Fable 5.1 that discount is steep enough that reading 2M tokens from the cache in the example session costs $0.50, just 5% of the session total.
Because reads are such a small slice, the cheap hit narrows the session gap only a little: $10.50 against $4.40, or 2.4x. The money goes to cache writes, which cost $6.50 on Fable 5.1 and make up 62% of its session. Of the $6.10 difference per session, $3.90 comes from writes and $1.50 from output. At 110 sessions a month, that is $1,155.00 against $484.00 at API rates.
When does Anthropic suggest Fable 5.1 over Opus 5.5?
Anthropic's guidance describes a default and an escalation. It calls Opus 5.5 its recommended starting model for most work and says, "It performs at the level of Claude Fable 5.1 on most work and costs 40% less to run than Opus 5." Fable 5.1 is Anthropic's most capable generally available model, aimed at demanding reasoning and long-horizon agentic coding, and Anthropic suggests it when Opus-tier results fall short.
That points to a working pattern rather than a single pick: run Opus 5.5 by default and move a task to Fable 5.1 when the result isn't good enough. In Claude Code, /model fable switches to it. Each model keeps its own cache in Claude Code, though, so the first request after a switch writes the whole context again at Fable 5.1's $12.50 or $20 per million.
The strength claims overlap. Anthropic lists long-running agentic coding and multistep research for Fable 5.1, and agentic coding, computer use, and codebase-wide migrations and audits for Opus 5.5. This page can't check either claim, but it can price them, and on identical token counts the premium for Fable 5.1 is 2.5x on every uncached token.
Effort, data retention, and fast mode
Both models accept 1M tokens of context at standard rates, write up to 128K tokens of output, and use Anthropic's newer tokenizer, which counts about 30% more tokens than earlier Claude models for the same text. A prompt is the same size on either, so the rate card alone sets the gap for input.
Output is less predictable, because the effort defaults differ. Fable 5.1 starts at high and Opus 5.5 at medium, and Anthropic notes that changing effort on Opus 5.5 keeps the prompt cache. Output already makes up 24% of the Fable 5.1 session in the example, and a higher effort setting tends to raise how much each request writes, at $50 per million.
Two facts can settle the question outright. Fable 5.1 requires 30-day data retention, so it is not available under zero data retention. And for latency-sensitive work, Opus 5.5 offers fast mode, a research preview on the Claude API at $8 input and $40 output per million, still below Fable 5.1's standard rates. You'll find both in Claude Code, Cursor, OpenRouter, OpenCode, and GitHub Copilot.
Prompt caching
How Anthropic bills cached tokens
Anthropic
Claude caches a prompt prefix up to a breakpoint. One top-level cache_control field places the breakpoint automatically and moves it as the conversation grows, or you can mark up to 4 blocks yourself. Claude Code manages caching for you.
A 5-minute cache write costs 1.25x the input price and a 1-hour write costs 2x. A cache hit costs 0.1x input on most models, 0.05x on Claude Opus 5.5, and 0.025x on Claude Fable 5.1 and Claude Mythos 5.1. Every hit restarts the cache lifetime at no charge.
Anthropic's rule of thumb: a 5-minute write pays for itself after one cache read, and a 1-hour write after two.
In Claude Code, the main conversation uses the 1-hour cache on a Claude subscription and the 5-minute cache with an API key. Each model has its own cache, so switching models starts over.
Source: Anthropic: Prompt caching
Your own numbers
See what Claude Fable 5.1 and Claude Opus 5.5 really cost you.
everyaitoken reads your Claude Code, Cursor, OpenCode, and OpenRouter history on your Mac and prices every request at API rates, with what caching saved or cost. $9 once.
FAQ
Questions
Is Claude Opus 5.5 as capable as Claude Fable 5.1?
Anthropic says Opus 5.5 performs at the level of Fable 5.1 on most work. It still calls Fable 5.1 its most capable generally available model and suggests it when Opus-tier results fall short. This page compares published prices and specs and doesn't test either model.
Why is the session 2.4x more expensive on Fable 5.1 when the rates are 2.5x?
Cache hits. Fable 5.1 bills a hit at 2.5% of its input price and Opus 5.5 at 5%, so reads cost $0.25 and $0.20 per million. Reads are a small part of the session, so the gap narrows only slightly.
Is Claude Fable 5.1 the default model in Claude Code?
No. Fable 5.1 is not the default on any plan, and you choose it with /model fable. Opus 5.5 is the default on Pro, Max, Team, and Enterprise plans and with an Anthropic API key.
How can I compare what the two models cost me?
Claude Code keeps a local log of the tokens each request used. EveryToken, a $9 one-time Mac app, prices that log at API rates by model and shows what the cache saved, so you can see what moving a share of your work to Fable 5.1 would add.
Sources
- Anthropic: Pricing
- Anthropic docs: Claude Fable 5.1
- Anthropic: Claude Fable 5.1 and Claude Mythos 5.1
- Claude Code docs: Model configuration
- Cursor docs: Claude Fable 5.1
- OpenRouter: Claude Fable 5.1
- OpenCode docs: Zen
- GitHub Docs: Supported AI models in Copilot
- Anthropic docs: Claude Opus 5.5
- Anthropic: Claude Opus 5.5
- Anthropic docs: Fast mode
- Cursor docs: Claude Opus 5.5
- OpenRouter: Claude Opus 5.5
- Anthropic: Prompt caching