Claude Fable 5 pricing
Anthropic's top model for long-running agentic work, priced at twice Opus 5. Worth it only where a longer autonomous run genuinely replaces human supervision, since the token cost compounds across a multi-hour session.
List prices checked
By Anthropic · Frontier tier
Token pricing
Per million tokens, from the official price list.
| Token type | Price per 1M | Notes |
|---|---|---|
| Input | $10.00 | Everything you send, including the system prompt |
| Output | $50.00 | Everything the model generates, reasoning tokens included |
| Cached input | $1.00 | Repeated prefixes served from cache |
| Blended (3:1) | $20.00 | One number for ranking, at three input tokens per output token |
Source: Anthropic pricing (August 2026). Published list prices only — committed-use discounts, regional variation, and per-request charges are not included. Last checked against the vendor’s page on 2026-08-07.
Capabilities
| Property | Value |
|---|---|
| Context window | 200K tokens |
| Max output | 64K tokens |
| Inputs accepted | text, image |
| Cost to fill the context once | $2.00 |
What it costs per month
Token prices applied to three workload shapes. Input-heavy and output-heavy workloads rank models very differently.
In-app chat assistant
A support or help assistant handling a few thousand conversations a month, with short prompts and short answers.
$450/mo
- 20M input tokens$200
- 5M output tokens$250
- If fully cached$270
RAG over documents
Retrieval-augmented answering where each request stuffs several retrieved passages into the prompt, so input dominates.
$3,750/mo
- 300M input tokens$3,000
- 15M output tokens$750
- If fully cached$1,050
Coding agent
An agentic loop that reads files and writes patches, generating heavy output and re-sending large context on every turn.
$9,000/mo
- 400M input tokens$4,000
- 100M output tokens$5,000
- If fully cached$5,400
Claude Fable 5 vs other frontier models
Same capability tier, ranked by blended token price.
| Model | Input | Output | Context | Compare |
|---|---|---|---|---|
| Claude Opus 5 | $5.00 | $25.00 | 200K | Head to head |
| Claude Sonnet 5 | $3.00 | $15.00 | 200K | Head to head |
| Grok 4 | $3.00 | $15.00 | 256K | Head to head |
| Cohere Command A | $2.50 | $10.00 | 256K | Head to head |
| OpenAI o3 | $2.00 | $8.00 | 200K | Head to head |
| Mistral Large | $2.00 | $6.00 | 131K | Head to head |
Questions about pricing
- How much does the Claude Fable 5 API cost in 2026?
- $10.00 per million input tokens and $50.00 per million output tokens, from Anthropic's published pricing. The table above converts that into a monthly figure for three common workloads.
- What is the Claude Fable 5 context window?
- 200K tokens, with up to 64K tokens of output per request. Filling the window costs whatever that many input tokens cost — at $10.00 per million, a single full-context request runs about $2.00.
- Does Claude Fable 5 support prompt caching?
- Yes. Cached input is billed at $1.00 per million tokens, roughly 10× cheaper than uncached input. For an agent that resends the same system prompt and file context on every turn, that discount usually matters more than the headline rate does.
- Is Claude Fable 5 cheaper than the alternatives?
- On a blended 3:1 input-to-output basis it works out at $20.00 per million tokens. The comparison table below ranks it against the other frontier models in the dataset, and the head-to-head pages model both against identical workloads.