Grok 4.3 pricing
Mid-range LLM from xAI —$1.25 input and$2.50 output per 1M tokens.
GAMid-rangeLong-context tier
- Input /1M
- $1.25
- Output /1M
- $2.50
- Cached /1M
- $0.20
- Blended
- $1.88
- Context
- 1M
API model id: grok-4.3
Full price matrix
| Rate | Input /1M | Output /1M | Notes |
|---|---|---|---|
| Standard | $1.25 | $2.50 | |
| Cached input (read) | $0.20 | — | Prompt-cache hit |
| Long context (> 200,000 tokens) | $2.50 | $5.00 | Input ×2, output ×2 |
List prices, USD, directional. Rates are provider list prices per 1M tokens and are meant for comparison, not billing. Batch rates shown as 50% off are derived where a provider offers batch but does not publish a separate figure. Preview, promo, intro, peak/off-peak, long-context, and third-party-host prices are labeled where they apply. Token counts vary by tokenizer, so per-token price is not always a like-for-like cost. Always confirm with the provider before relying on a number. Prices as of 2026-08-13.
Source: https://docs.x.ai/developers/pricing
Pricing notes
- Long-context pricing: above 200,000 tokens, input is billed at ×2 ($2.50) and output at ×2 ($5.00). See long-context pricing.
- Once a prompt reaches the 200K-token long-context threshold the whole request bills long-context rates: $2.50 input / $5 output, cached input $0.40. xAI's Batch API gives this model a 20% discount.