GPT-6 Astra pricing
Flagship LLM from OpenAI —$10.00 input and$50.00 output per 1M tokens.
GAFlagshipLong-context tier
- Input /1M
- $10.00
- Output /1M
- $50.00
- Cached /1M
- $1.00
- Blended
- $30.00
- Context
- 1.05M
- Max output
- 128K
API model id: gpt-6-astra
Full price matrix
| Rate | Input /1M | Output /1M | Notes |
|---|---|---|---|
| Standard | $10.00 | $50.00 | |
| Cached input (read) | $1.00 | — | Prompt-cache hit |
| Cache write (5 min) | $12.50 | — | One-time write to store a prefix |
| Batch API | $5.00 | $25.00 | Async, 50% off |
| Long context (> 272,000 tokens) | $20.00 | $75.00 | Input ×2, output ×1.5 |
List prices, USD, directional. Rates are provider list prices per 1M tokens and are meant for comparison, not billing. Batch rates shown as 50% off are derived where a provider offers batch but does not publish a separate figure. Preview, promo, intro, peak/off-peak, long-context, and third-party-host prices are labeled where they apply. Token counts vary by tokenizer, so per-token price is not always a like-for-like cost. Always confirm with the provider before relying on a number. Prices as of 2026-09-04.
Source: https://developers.openai.com/api/docs/pricing
Pricing notes
- Long-context pricing: above 272,000 tokens, input is billed at ×2 ($20.00) and output at ×1.5 ($75.00). See long-context pricing.
- Launched 2026-09-03; access was staged at launch (limited organizations first, then broader API and cloud-marketplace availability) — confirm your own access before planning against this rate. Long context (past ~272K tokens) bills the whole request at $20 input / $75 output, cached input $2.00. A fast mode is offered at 2x the price for 2x the speed.
Similar flagship models
More from OpenAI
GPT-5 · $1.25/$10.00GPT-5.6 Sol · $4.00/$20.00GPT-5.5 · $5.00/$30.00o3-pro · $20.00/$80.00GPT-5.5 Pro · $30.00/$180.00o4-mini · $1.10/$4.40o3 · $2.00/$8.00GPT-5.6 Terra · $2.00/$12.00GPT-5.4 · $2.50/$15.00GPT-5 nano · $0.05/$0.40GPT-5.6 Luna · $0.20/$1.20GPT-5.4 nano · $0.20/$1.25GPT-5.4 mini · $0.75/$4.50