Gemini 2.5 Flash Pricing
Gemini 2.5 Flash is a mid-tier language model from Google, priced at $0.30 per 1M input tokens and $2.50 per 1M output tokens.
What does Gemini 2.5 Flash cost in practice?
These are real-world monthly estimates based on common workload sizes. They assume a 2:1 input-to-output ratio (typical of chat + retrieval-augmented apps).
| Workload | Input / mo | Output / mo | Monthly cost | Annualized |
|---|---|---|---|---|
| Light prototyping | 0.1M tokens | 0.1M tokens | $0.15 | $1.86 |
| Medium production | 5.0M tokens | 2.5M tokens | $7.75 | $93.00 |
| Heavy production | 50.0M tokens | 25.0M tokens | $77.50 | $930.00 |
Need to model your own usage? Use the LLM pricing calculator →
Gemini 2.5 Flash features
- Vision
- Tool Use
- Streaming
- Reasoning
- Audio
Alternatives to Gemini 2.5 Flash
Cheaper, same-tier alternatives
If Gemini 2.5 Flash is more than you need, these mid-tier models cost less per token:
Need the cheapest Google option? Gemini 2.0 Flash starts at $ 0.10 input / $ 0.40 output per 1M tokens →
FAQ
How much does Gemini 2.5 Flash cost per 1M tokens?
Gemini 2.5 Flash costs $0.30 per 1M input tokens and $2.50 per 1M output tokens. Cached input drops to $0.03 per 1M tokens — a 90% discount on repeated context.
What is Gemini 2.5 Flash's context window?
Gemini 2.5 Flash supports a 1,000,000-token context window with a maximum output of 64,000 tokens per response.
What features does Gemini 2.5 Flash support?
Vision, Tool Use, Streaming, Reasoning, Audio.
How much does Gemini 2.5 Flash cost for medium production usage?
For ~5M input + 2.5M output tokens per month, Gemini 2.5 Flash costs approximately $7.75/mo (≈$ 93.00/yr).
What are cheaper alternatives to Gemini 2.5 Flash?
Same-tier alternatives that cost less: Claude Haiku 4.5 , GPT-4.1 Mini , GPT-5 Mini .