Gemini pricing at a glance
| Model | Input $/1M | Output $/1M | Best for |
|---|---|---|---|
| Gemini 3.1 Flash-Lite | $0.25 | $1.50 | Cheapest, high-volume tasks |
| Gemini 3.5 Flash | $1.50 | $9.00 | Balanced speed/cost |
| Gemini 3.6 Flash | $1.50 | $7.50 | Latest Flash release |
| Gemini 3.1 Pro | $2.00 | $12.00 | Highest-capability tasks |
See Google's official Gemini pricing page →
Frequently asked questions
How much does the Gemini API cost?
Flash-Lite is $0.25/$1.50 per million input/output tokens, Flash is roughly $1.50/$7.50-$9.00, and Pro is $2.00/$12.00. Rates change often — verify on Google's Gemini pricing page.
Is Gemini cheaper than Claude or GPT?
Gemini's cheapest tier is usually the lowest-cost option for simple tasks, but the ranking depends on your input/output ratio and caching/batch usage. Compare directly with the calculator above.
Does Gemini support caching and batch discounts?
Yes — Google's Batch API is about 50% off, and context caching can save up to roughly 90% on cached input tokens. Toggle both in the calculator to see the effect on your bill.
Compare against Claude and GPT
Want the full picture across providers? Use the universal LLM cost calculator to compare Gemini against Anthropic's Claude models and OpenAI's GPT models side by side for the same workload.