Claude Opus 5 vs o3 pricing
Side-by-side LLM API pricing.o3 is cheaper on input ($2.00 vs $5.00 /1M), and o3 is cheaper on output ($8.00 vs $25.00 /1M). On a blended average, o3 is lower — but which wins for you depends on your input/output mix, so check the example workloads below or run the calculator.
| Attribute | Claude Opus 5 | o3 |
|---|---|---|
| Provider | Anthropic | OpenAI |
| Tier | Flagship | Reasoning |
| Status | GA | GA |
| Context window | 1M | 200K |
| Input /1M | $5.00 | $2.00 |
| Output /1M | $25.00 | $8.00 |
| Cached input /1M | $0.50 | $0.50 |
| Batch input /1M | $2.50 | $1.00 |
| Batch output /1M | $12.50 | $4.00 |
| Blended (avg in+out) | $15.00 | $5.00 |
| 1M in + 1M out | $30.00 | $10.00 |
| 1M in (90% cached) + 100K out | $3.45 | $1.45 |
List prices, USD, directional. Rates are provider list prices per 1M tokens and are meant for comparison, not billing. Batch rates shown as 50% off are derived where a provider offers batch but does not publish a separate figure. Preview, promo, intro, peak/off-peak, long-context, and third-party-host prices are labeled where they apply. Token counts vary by tokenizer, so per-token price is not always a like-for-like cost. Always confirm with the provider before relying on a number. Prices as of 2026-08-03.
Tokenizer caveat: Claude Opus 5 uses a tokenizer that produces more tokens per unit of text, so identical content is billed as a different number of tokens on each model. Per-token price is therefore not a like-for-like cost — compare cost per task, not just the sticker rate.
Which should you pick?
For a balanced job (1M input + 1M output), o3 costs $10.00 versus $30.00. If your prompts are large and reused, prompt caching changes the maths — the cache-heavy row above shows a 90%-cached input scenario. If your work can run asynchronously, both providers' batch rates cut the bill further where offered.
Full price matrices: Claude Opus 5 pricing → · o3 pricing → · Back to the LLM pricing hub →