o3 pricing
Reasoning LLM from OpenAI —$2.00 input and$8.00 output per 1M tokens.
GAReasoning
- Input /1M
- $2.00
- Output /1M
- $8.00
- Cached /1M
- $0.50
- Blended
- $5.00
- Context
- 200K
- Max output
- 100K
API model id: o3
Full price matrix
| Rate | Input /1M | Output /1M | Notes |
|---|---|---|---|
| Standard | $2.00 | $8.00 | |
| Cached input (read) | $0.50 | — | Prompt-cache hit |
| Batch API | $1.00 | $4.00 | Async, 50% off |
List prices, USD, directional. Rates are provider list prices per 1M tokens and are meant for comparison, not billing. Batch rates shown as 50% off are derived where a provider offers batch but does not publish a separate figure. Preview, promo, intro, peak/off-peak, long-context, and third-party-host prices are labeled where they apply. Token counts vary by tokenizer, so per-token price is not always a like-for-like cost. Always confirm with the provider before relying on a number. Prices as of 2026-08-03.
Source: https://developers.openai.com/api/docs/pricing