Verified 2026-10-07 · sourced from Anthropic
100K Claude Fable 5.1 Tokens — Cost Breakdown
Use this guide to benchmark budgets for 100,000 tokens. Standard pricing is $10.00 per million input tokens and $50.00 per million output tokens. Cached input, when available, reduces costs significantly for repeated contexts.
Reference rates: Standard · global · no long-context premium
USD per million text tokens. Verified 2026-10-07 · Official source
| Mode / prompt size | Input | Cache read | Cache write | Output |
|---|---|---|---|---|
| Standard | $10 | $0.25 | $12.5 (5m) / $20 (1h) | $50 |
| Batch | $5 | Not estimated | Not estimated | $25 |
1,000,000 context tokens · max output 128,000. Local token counts are estimates; use the provider’s usage counts for an invoice estimate.
Text-only estimate. Local token counts are approximate; use provider-reported usage for billing. Tool, image, audio and video charges are excluded.
No long-context premium within the 1M context window. Cache writes: 5-minute and 1-hour retention are separate input billing operations.
US-only inference adds 10% to token rates and is excluded from this global-endpoint estimate. Newer Claude tokenization differs from the local estimator; no fixed conversion factor is applied.
Billed output includes reasoning/thinking tokens that cannot be inferred from pasted final text; use API-reported output usage for accurate billing.
Batch: Batch cache read/write combinations are not estimated here; use uncached input or consult the official source.
Scenario breakdown
Cost estimates for different input/output distributions using 100K total tokens.
| Scenario | Tokens in | Tokens out | Standard cost | Cached cost |
|---|---|---|---|---|
Balanced conversation 50% input · 50% output | 50,000 | 50,000 | $3.00 | $2.51 |
Input-heavy workflow 80% input · 20% output | 80,000 | 20,000 | $1.80 | $1.02 |
Generation heavy 30% input · 70% output | 30,000 | 70,000 | $3.80 | $3.51 |
Cached system prompt 90% cached input · 10% fresh output | 90,000 | 10,000 | $1.40 | $0.522 |
Workload multipliers
Convert 100K tokens into daily and monthly run-rate budgets.
| Profile | Runs/day | Tokens/day | Daily cost | Monthly cost | Cached daily | Cached monthly |
|---|---|---|---|---|---|---|
| Single workload | 1 | 100,000 | $3.00 | $90.00 | $0.522 | $15.67 |
| Daily batch (10 runs) | 10 | 1,000,000 | $30.00 | $900.00 | $5.22 | $156.75 |
| Team workload (100 runs) | 100 | 10,000,000 | $300.00 | $9000.00 | $52.25 | $1567.50 |
Frequently asked questions
What is the standard cost of 100K Claude Fable 5.1 tokens?
100K tokens in a 50/50 conversation mix cost roughly $3.00 at the published Anthropic rates.
What happens if cached input is enabled?
With cached contexts, the same 100K tokens drop to $2.51 because input costs fall to $0.250 per million tokens.
How many requests does 100K tokens cover?
If your prompts average around 3,000 tokens per call, 100K total tokens cover about 33 requests.
How fresh is the pricing information?
Prices are taken from https://platform.claude.com/docs/en/about-claude/pricing and were last verified on 2026-10-07. models.json keeps this guide in sync with upstream changes.