Claude Opus 5.5 pricing
Flagship LLM from Anthropic —$4.00 input and$20.00 output per 1M tokens.
GAFlagship
- Input /1M
- $4.00
- Output /1M
- $20.00
- Cached /1M
- $0.20
- Blended
- $12.00
- Context
- 1M
- Max output
- 128K
API model id: claude-opus-5-5
Full price matrix
| Rate | Input /1M | Output /1M | Notes |
|---|---|---|---|
| Standard | $4.00 | $20.00 | |
| Cached input (read) | $0.20 | — | Prompt-cache hit |
| Cache write (5 min) | $5.00 | — | One-time write to store a prefix |
| Batch API | $2.00 | $10.00 | Async, 50% off |
List prices, USD, directional. Rates are provider list prices per 1M tokens and are meant for comparison, not billing. Batch rates shown as 50% off are derived where a provider offers batch but does not publish a separate figure. Preview, promo, intro, peak/off-peak, long-context, and third-party-host prices are labeled where they apply. Token counts vary by tokenizer, so per-token price is not always a like-for-like cost. Always confirm with the provider before relying on a number. Prices as of 2026-09-29.
Source: https://platform.claude.com/docs/en/about-claude/pricing
Pricing notes
- Tokenizer: Newer tokenizer produces roughly 30% more tokens per unit of text — per-token price is not directly comparable across tokenizer generations. See tokenizer.
- Anthropic's default flagship model, launched 2026-09-22, replacing Claude Opus 5. List price is 20% below Opus 5 ($4 / $20 vs $5 / $25) and cache reads are priced at 0.05x input ($0.20) instead of the usual 0.1x; Anthropic estimates roughly 40% lower cost than Opus 5 on a typical workload. The 1M-token context window is billed at standard rates. Fast mode (first-party API only) is priced at $8 input / $40 output.