o4-mini pricing
Reasoning LLM from OpenAI —$1.10 input and$4.40 output per 1M tokens.
GAReasoning
- Input /1M
- $1.10
- Output /1M
- $4.40
- Cached /1M
- $0.275
- Blended
- $2.75
- Context
- 200K
API model id: o4-mini
Full price matrix
| Rate | Input /1M | Output /1M | Notes |
|---|---|---|---|
| Standard | $1.10 | $4.40 | |
| Cached input (read) | $0.275 | — | Prompt-cache hit |
| Batch API | $0.55 | $2.20 | Async, 50% off |
List prices, USD, directional. Rates are provider list prices per 1M tokens and are meant for comparison, not billing. Batch rates shown as 50% off are derived where a provider offers batch but does not publish a separate figure. Preview, promo, intro, peak/off-peak, long-context, and third-party-host prices are labeled where they apply. Token counts vary by tokenizer, so per-token price is not always a like-for-like cost. Always confirm with the provider before relying on a number. Prices as of 2026-08-03.
Source: https://developers.openai.com/api/docs/pricing