Claude Haiku 5.5 Pricing & Token Costs (2026)
Per 1M tokens: input $0.10 · output $0.50 · cached $0.010. Context window 1,000,000 tokens with source and verification details.
Reference rates: Standard · ≤100K input · global
USD per million text tokens. Verified 2026-10-08 · Official source
| Mode / prompt size | Input | Cache read | Cache write | Output |
|---|---|---|---|---|
| Standard ≤100,000 input | $0.1 | $0.01 | $0.125 (5m) / $0.2 (1h) | $0.5 |
| Standard >100,000 input | $0.5 | $0.05 | $0.625 (5m) / $1 (1h) | $2.5 |
| Batch ≤100,000 input | $0.05 | $0.005 | $0.0625 (5m) / $0.1 (1h) | $0.25 |
| Batch >100,000 input | $0.25 | $0.025 | $0.3125 (5m) / $0.5 (1h) | $1.25 |
1,000,000 context tokens · max output 128,000. Local token counts are estimates; use the provider’s usage counts for an invoice estimate.
Text-only estimate. Local token counts are approximate; use provider-reported usage for billing. Tool, image, audio and video charges are excluded.
Up to 100,000 input tokens uses the lower tier. Above 100,000, the full request uses higher input, cache and output rates, not just the excess tokens. The threshold is per request; daily call volume does not change it.
For prompt-size tier selection, total input includes uncached, cache-read and cache-creation tokens, using the provider’s total-input definition. This calculator applies one selected input billing operation to all input; it does not split mixed cache usage.
US-only inference adds 10% to token rates and is excluded from this global-endpoint estimate. Newer Claude tokenization differs from the local estimator; no fixed conversion factor is applied.
Billed output includes reasoning/thinking tokens that cannot be inferred from pasted final text; use API-reported output usage for accurate billing.
Batch: Batch halves token rates and stacks with cache read/write pricing. Cache hits in asynchronous batches are best effort, not guaranteed. Cache writes use separate 5-minute and 1-hour retention rates.
TL;DR — Pricing Quick Summary
- ✓Input pricing: $0.10 per 1M tokens ($0.0001 per 1K)
- ✓Output pricing: $0.50 per 1M tokens ($0.0005 per 1K)
- ✓Prompt caching: $0.010 per 1M tokens — save 90% on repeated context
- ✓Context window: 1,000,000 tokens
- ✓Typical monthly cost: $2.00 for 10M input + 2M output across short requests; Standard · ≤100K input · global
- ✓Daily cost example: $0.0300 for 100K tokens (50K in, 50K out)
Key metrics
- Context window
- 1,000,000 tokens
- Input price
- $0.10 / 1M tokens
- Output price
- $0.50 / 1M tokens
- Cached input
- $0.010 / 1M tokens
Official link:https://platform.claude.com/docs/en/about-claude/pricing
Last verified: 2026-10-08
- · Text-only estimate. Local token counts are approximate; use provider-reported usage for billing. Tool, image, audio and video charges are excluded.
- · Up to 100,000 input tokens uses the lower tier. Above 100,000, the full request uses higher input, cache and output rates, not just the excess tokens. The threshold is per request; daily call volume does not change it.
- · For prompt-size tier selection, total input includes uncached, cache-read and cache-creation tokens, using the provider’s total-input definition. This calculator applies one selected input billing operation to all input; it does not split mixed cache usage.
- · US-only inference adds 10% to token rates and is excluded from this global-endpoint estimate. Newer Claude tokenization differs from the local estimator; no fixed conversion factor is applied.
- · Billed output includes reasoning/thinking tokens that cannot be inferred from pasted final text; use API-reported output usage for accurate billing.
Multi-currency (per 1M tokens)
| Currency | Input | Cached | Output |
|---|---|---|---|
| USD | $0.10 | $0.01 | $0.50 |
| CNY | ¥0.72 | ¥0.07 | ¥3.58 |
| EUR | 0,09 € | 0,01 € | 0,46 € |
| JPY | ¥15 | ¥1 | ¥73 |
* Live search cost uses sources / 1000 × price and currently applies to xAI Grok only.
Frequently Asked Questions
What is the cost per 1M tokens for Anthropic Claude Haiku 5.5?
Anthropic Claude Haiku 5.5 costs $0.10 per 1M input tokens and $0.50 per 1M output tokens, with cached input at $0.010 per 1M tokens.
How much does it cost per 1K tokens?
Per 1K tokens: $0.0001 for input and $0.0005 for output. This is useful for calculating costs for smaller workloads or individual API calls.
What is the estimated monthly cost for typical usage?
For 10M input + 2M output tokens per month across requests within the short-context tier, Anthropic Claude Haiku 5.5 would cost approximately $2.00. Daily usage of 100K tokens (50K in, 50K out) costs about $0.0300.
Does Anthropic Claude Haiku 5.5 offer a free tier?
Check Anthropic's official documentation for free tier availability. Some providers offer free credits for new users or limited free usage. Visit https://platform.claude.com/docs/en/about-claude/pricing for current free tier details.
How does prompt caching work to reduce costs?
With prompt caching enabled, input pricing drops to $0.010 per 1M tokens for repeated context (a 90% discount), while output remains $0.50 per 1M tokens. Caching is ideal for repeated prompts or system messages.
What is the context window size for Claude Haiku 5.5?
Claude Haiku 5.5 supports a 1,000,000 token context window. This determines the maximum combined length of your input prompt and output response.
How frequently is this pricing information updated?
All prices reference official Anthropic documentation (https://platform.claude.com/docs/en/about-claude/pricing), last verified on 2026-10-08. Review the source before budgeting; other catalog entries may have older verification dates.
How can I calculate exact costs for my use case?
Use our free token calculator to estimate costs based on your specific usage pattern. The calculator supports all major models and shows costs in multiple currencies. You can also compare costs across different models to find the most economical option.