Gemini 3.6 Flash pricing
Mid-range LLM from Google —$0.75 input and$3.75 output per 1M tokens.
GAMid-rangeIntro price until 2026-12-31Promo available
- Input /1M
- $0.75
- Output /1M
- $3.75
- Cached /1M
- $0.075
- Blended
- $2.25
- Context
- 1M
API model id: gemini-3.6-flash
Full price matrix
| Rate | Input /1M | Output /1M | Notes |
|---|---|---|---|
| Standard | $0.75 | $3.75 | |
| Cached input (read) | $0.075 | — | Prompt-cache hit |
| Batch API | $0.375 | $1.88 | Derived: 50% off (async) |
| Standard (after 2026-12-31) | $1.50 | $7.50 | Intro pricing ends |
List prices, USD, directional. Rates are provider list prices per 1M tokens and are meant for comparison, not billing. Batch rates shown as 50% off are derived where a provider offers batch but does not publish a separate figure. Preview, promo, intro, peak/off-peak, long-context, and third-party-host prices are labeled where they apply. Token counts vary by tokenizer, so per-token price is not always a like-for-like cost. Always confirm with the provider before relying on a number. Prices as of 2026-09-04.
Source: https://ai.google.dev/gemini-api/docs/pricing
Pricing notes
- Intro pricing: the $0.75/$3.75 rate is introductory until 2026-12-31; the standard rate afterwards is $1.50 input / $7.50 output.
- Promotion: Introductory pricing through 2026-12-31; from 2027-01-01 input rises to $1.50, output to $7.50 and cached input to $0.15..
- Cache storage billed separately at $0.50 per 1M tokens per hour ($1.00 from 2027-01-01).