Free cloud & LLM pricing tools — AWS EC2 + LLM API Pricing Explorer

GPT-6 Luna vs Gemini 2.5 Flash-Lite pricing

Side-by-side LLM API pricing.Input rates are equal ($0.10 vs $0.10 /1M), and Gemini 2.5 Flash-Lite is cheaper on output ($0.40 vs $0.50 /1M). On a blended average, Gemini 2.5 Flash-Lite is lower — but which wins for you depends on your input/output mix, so check the example workloads below or run the calculator.

AttributeGPT-6 LunaGemini 2.5 Flash-Lite
ProviderOpenAIGoogle
TierBudgetBudget
StatusGAGA
Context window1.05M1M
Input /1M$0.10$0.10
Output /1M$0.50$0.40
Cached input /1M$0.01$0.01
Batch input /1M$0.05$0.05
Batch output /1M$0.25$0.20
Blended (avg in+out)$0.30$0.25
1M in + 1M out$0.6000$0.5000
1M in (90% cached) + 100K out$0.0690$0.0590

List prices, USD, directional. Rates are provider list prices per 1M tokens and are meant for comparison, not billing. Batch rates shown as 50% off are derived where a provider offers batch but does not publish a separate figure. Preview, promo, intro, peak/off-peak, long-context, and third-party-host prices are labeled where they apply. Token counts vary by tokenizer, so per-token price is not always a like-for-like cost. Always confirm with the provider before relying on a number. Prices as of 2026-09-29.

Which should you pick?

For a balanced job (1M input + 1M output), Gemini 2.5 Flash-Lite costs $0.5000 versus $0.6000. If your prompts are large and reused, prompt caching changes the maths — the cache-heavy row above shows a 90%-cached input scenario. If your work can run asynchronously, both providers' batch rates cut the bill further where offered.

Full price matrices: GPT-6 Luna pricing → · Gemini 2.5 Flash-Lite pricing → · Back to the LLM pricing hub →