cd /news/large-language-models/fable-5-1-costs-twice-as-much-as-opu… · home › topics › large-language-models › article
[ARTICLE · art-140701] src=dev.to ↗ pub= topic=large-language-models verified=true sentiment=· neutral

Fable 5.1 costs twice as much as Opus 5, until your cache gets big enough

A developer built llmabacus.com, an LLM API price table, and used it to calculate the exact cache-read crossover point between Claude Fable 5.1 and Opus 5. Fable 5.1 halves cache-read pricing to $0.25 per million tokens while keeping $10/$50 input/output rates, so it only becomes cheaper than Opus 5 past roughly 140K cached tokens per request — and cache-write premiums may prevent heavy agent users from ever reaching that point.

by read3 min views1 publishedSep 28, 2026

Claude Fable 5.1 shipped on September 1 with the same $10 input / $50 output per million tokens as Fable 5. One number moved: cache reads dropped from $1 to $0.25 per million.

On the price sheet it still costs twice what Opus 5 does ($5 / $25). But look at the cache-read column and the order flips. Fable 5.1 charges $0.25 there, Opus 5 charges $0.50. So which one is cheaper depends on how much of each request is served from cache, and I wanted the actual crossover instead of a vibe.

Cached tokens times the cache-read price, fresh input times the input price, output times the output price. Fable 5.1 only changed the first term, so the saving over Fable 5 is exactly as big as that term's share of your bill.

Two workloads I priced out from the table on my site:

Per request Fable 5 Fable 5.1 Opus 5
Agent: 100K cached, 2K new input, 1K output $0.170 $0.095 $0.085
Chat: 5K cached, 1K new input, 800 output $0.055 $0.051 $0.028

In the agent loop, cache reads are about 60% of the Fable 5 bill, so the price cut takes 44% off each request. That lines up with the "around 45% for heavy agent use" figure in the launch coverage.

The chat case barely moves. Output dominates, and Opus 5 stays roughly half the price. Switching there is just paying more.

Fable 5.1 pays double for fresh input and output, and half for cache reads. Set the two bills equal:

0.25 * R = 5 * U + 25 * O
R = 20 * U + 100 * O

R is cached tokens per request, U is fresh input, O is output. Plug in the agent numbers above and the line sits at 140K cached tokens. Both models cost $0.105 a request there.

Past that point the gap keeps widening. At 200K cached, Fable 5.1 is about 11% cheaper. Double the cache again and it's more than a quarter cheaper.

Output length is what pushes the line out. Every extra thousand output tokens needs another hundred thousand cached tokens before Fable 5.1 catches up.

My first pass stopped there and the conclusion was "long-context agents should move to Fable 5.1." Then I remembered that tokens have to get into the cache before they can be read from it.

Anthropic bills a 5-minute cache write at 1.25x the input price for Opus 5. My price table has no separate write price for Fable 5.1. If it follows the same 1.25x rule, writing a 200K prefix costs $1.25 more on Fable 5.1 than on Opus 5. At 200K cached, Fable 5.1 saves $0.015 per request. That's more than 80 requests on the same prefix before the write premium is paid back.

So if your prefix expires every five minutes, or your agent keeps rewriting its context, you may never reach the crossover at all. Measure how many requests actually reuse each prefix before you switch.

Pull three averages from your logs: cached tokens, fresh input and output per request. Then check how many requests reuse the same prefix.

20U + 100O cached tokens: stay on Opus 5. I built llmabacus.com, the LLM API price table these numbers come from. The full write-up, in Chinese, is at https://www.llmabacus.com/articles/fable-5-1-cache-read-vs-opus-5

── more in #large-language-models 4 stories · sorted by recency
── more on @anthropic 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
→ Live at https://your-agent.zahid.host ✓
Get free account → Pricing
from €0/mo · no card required
LIVE [news/fable-5-1-costs-twic…] indexed:0 read:3min 2026-09-28 · —