Free cloud & LLM pricing tools — AWS EC2 + LLM API Pricing Explorer

Google API pricing

Google’s Gemini API pairs 1M-token context windows with some of the lowest budget-tier prices. Pro models are tier-priced by prompt size (a surcharge above 200K tokens); Flash and Flash-Lite are flat, with a 50%-off batch mode.

8 models below, cheapest on a blended average is Gemini 2.5 Flash-Lite at $0.25 /1M. Official pricing: https://ai.google.dev/gemini-api/docs/pricing.

ModelTierStatusContextInput /1MOutput /1MCached /1MBatch in /1M
Gemini 2.5 ProFlagshipGA1M$1.25$10.00$0.125$0.625
Gemini 3.1 PropreviewFlagshipPreview1M$2.00$12.00$0.20$1.00
Gemini 2.5 FlashMid-rangeGA1M$0.30$2.50$0.03$0.15
Gemini 3.6 FlashMid-rangeGA1M$1.50$7.50$0.15$0.75
Gemini 3.5 FlashMid-rangeGA1M$1.50$9.00$0.15$0.75
Gemini 2.5 Flash-LiteBudgetGA1M$0.10$0.40$0.01$0.05
Gemini 3.1 Flash-LiteBudgetGA1M$0.25$1.50$0.025$0.125
Gemini 3.5 Flash-LiteBudgetGA1M$0.30$2.50$0.03$0.15

List prices, USD, directional. Rates are provider list prices per 1M tokens and are meant for comparison, not billing. Batch rates shown as 50% off are derived where a provider offers batch but does not publish a separate figure. Preview, promo, intro, peak/off-peak, long-context, and third-party-host prices are labeled where they apply. Token counts vary by tokenizer, so per-token price is not always a like-for-like cost. Always confirm with the provider before relying on a number. Prices as of 2026-08-03.
Source: https://ai.google.dev/gemini-api/docs/pricing

Other providers

AnthropicOpenAIxAIDeepSeekMistralMeta (Llama)Alibaba (Qwen)

← Back to the LLM pricing hub · Pricing glossary