Last verified
MYTHOS-CLASS1M CONTEXTCACHE READ 0.025xTHINKING ALWAYS ONTEXT + VISION

Claude Fable 5.1 API Pricing

Claude Fable 5.1 shipped September 1, 2026 at exactly the rates of the model it replaces - $10/M input, $50/M output, $12.50/M and $20/M cache writes. One number moved: the cache hit fell from $1/M to $0.25/M. Anthropic prices cache reads at 0.1x base input on every other Claude model and at 0.025x on this one, and says so in a footnote on its own pricing page.

Input - per 1M tokens
$10.00/M
Same as Fable 5 unchanged
Output - per 1M tokens
$50.00/M
Model ID claude-fable-5-1 unchanged
Cached input - per 1M tokens
$0.25/M
Was $1 on Fable 5 -97.5%
Effective - agentic blend
$6.26/M
92/8 split - 82% cache
§ 01 / TERMINAL

Run the numbers.

Live calculator pre-loaded with Claude Fable 5.1's rates. Move the cache slider first: on this model the gap between a cache read and a cache write is 50x, the widest on any Claude row, so the read-to-write ratio decides the bill far more than the headline rates do.

$ /mo
Workload split
Prompt cache hit rate
Tokens you can process
Words equivalent (English)
Effective rate
Open full calculator (all models · share URL · CSV) →
§ 02 / SCENARIOS

Real-world presets.

§ 03 / TOKENIZER

Paste text. See tokens. See cost.

Estimate · anthropic-bpe-estimate · ≈2.6 chars/token Auto-counts as you type

This is a chars-per-token approximation, not a real tokenizer. Actual tokens vary by language, code density, and tool-call overhead — counts are typically ±10–20% off for English prose, more for code or non-Latin scripts. For exact billing, use the vendor's official tokenizer.

Characters 784
Words 129
Tokens (estimated) 302 tokens
Cost as input · uncached $0.00302 USD
Cost as output · uncached $0.0151 USD
Cost as cached input $0.00008 USD
§ 04 / SHELF

Up against the shelf.

All models →
Model Input /M Output /M Effective blended Context Best for
Claude Fable 5.1 Current $10.00 cache $0.25 $50.00 $6.26 agentic 92/8 1M Demanding reasoning and long-horizon agents
Claude Mythos 5.1 $10.00 cache $0.25 $50.00 $6.26 same card, restricted access 1M Glasswing cyber / bio access
Claude Fable 5 $10.00 cache $1.00 $50.00 $6.82 same base, 4x the cache read 1M The row this one replaces
Claude Opus 5 $5.00 cache $0.50 $25.00 $3.41 exactly half the card 1M Anthropic's general frontier tier
Claude Sonnet 5 $2.00 cache $0.20 $10.00 $1.36 speed / intelligence balance 1M Most production workloads
Claude Haiku 4.5 $1.00 cache $0.10 $5.00 $0.682 cheap high-volume tier 200K Support and classification
GPT-5.5 $5.00 cache $0.50 $30.00 $3.61 frontier competitor 1M OpenAI frontier agents

Frequently asked.

What the 75% cache cut is actually worth, why the headline saving and our blended figure disagree, and how 5.1 sits against Fable 5 and Opus 5.

Q · 01 What is Claude Fable 5.1 priced at? +
Anthropic's pricing table lists $10/M base input, $12.50/M 5-minute cache writes, $20/M 1-hour cache writes, $0.25/M cache hits, and $50/M output. Batch is half of base at $5/M and $25/M. Under this site's 92/8 agentic blend at an 82% cache-hit rate the effective rate is $6.2586/M.
Q · 02 What actually changed from Claude Fable 5? +
One number. Base input, both cache-write tiers, output and the batch rates are byte-for-byte identical to Fable 5. The cache hit went from $1/M to $0.25/M. Anthropic prices cache reads at 0.1x base input on every other Claude model; on Fable 5.1 and Mythos 5.1 it is 0.025x, and the pricing page carries a footnote saying precisely that.
Q · 03 Anthropic says up to 45% cheaper. Why does this page say 8%? +
Because the two figures measure different workloads, and both are correct. Our standard blend is 92% input / 8% output at an 82% cache-hit rate, and it charges every cache miss at the $12.50/M write rate. That leaves the effective rate at $6.2586 against Fable 5's $6.8244 - a 8.3% saving. Push the cache-hit rate to a theoretical 100% and our blend still only reaches 14%, because output is unchanged and contributes $4 of the total on its own. Anthropic's 45% therefore describes a much more input-heavy mix than 92/8, with many reads per write. If your agent re-reads one long cached context hundreds of times per write, you will land closer to the vendor's number than to ours.
Q · 04 So when is the cache cut worth the most? +
When input dominates and the same cached prefix is read many times. The gap between a read ($0.25) and a 5-minute write ($12.50) is 50x on this model - the widest ratio on any Claude row - so what decides your bill is not the hit rate alone but how many reads each write earns. A long-running agent on one stable context is the best case; short unrelated requests that each write a fresh cache are the worst, and there Fable 5.1 costs the same as Fable 5.
Q · 05 How does it compare with Claude Opus 5? +
Opus 5 is exactly half the card at $5/$25 and blends to $3.4122/M, so Fable 5.1 costs 1.8x as much on our assumptions. Anthropic's own guidance is to start with Opus 5 for most workloads and reach for Fable 5.1 for demanding reasoning and long-horizon agentic work, or when evals on Opus 5 at higher effort still fall short.
Q · 06 What is the difference between Fable 5.1 and Mythos 5.1? +
The same split as the 5.0 pair: identical rate cards, different access. Fable 5.1 is generally available with Anthropic's production safeguards. Mythos 5.1 is limited availability through the Glasswing programme for vetted cybersecurity and life-sciences organisations, with no self-serve API.
Q · 07 Is a per-token rate comparable across Claude generations? +
Not directly. Anthropic notes that Claude 4.7 and later models use a newer tokenizer producing roughly 30% more tokens for the same text, with the exact increase depending on content. A model on the newer tokenizer at the same nominal $/M is more expensive per page of real text than one on the older tokenizer. Anthropic puts 1M tokens at about 555k words on the current tokenizer against about 750k words before it.
Q · 08 Does the 1M context cost extra? +
No. Anthropic's pricing page states that Claude 4.6 and later models include the full 1M-token context window at standard pricing - a 900k-token request is billed at the same per-token rate as a 9k-token one - and that caching and batch discounts apply at standard rates across the whole window. Pinning inference to the US with inference_geo is the modifier that does cost more: a 1.1x multiplier on every token category.