Google

Gemini 2.0 Flash Pricing

Gemini 2.0 Flash is a budget-tier language model from Google, priced at $0.10 per 1M input tokens and $0.40 per 1M output tokens.

Input
$0.10
per 1M tokens
Output
$0.40
per 1M tokens
Cached input
$0.01
per 1M tokens 90% off
Context
1000K
tokens
Max output
8K
tokens

What does Gemini 2.0 Flash cost in practice?

These are real-world monthly estimates based on common workload sizes. They assume a 2:1 input-to-output ratio (typical of chat + retrieval-augmented apps).

Workload Input / mo Output / mo Monthly cost Annualized
Light prototyping 0.1M tokens 0.1M tokens $0.03 $0.36
Medium production 5.0M tokens 2.5M tokens $1.50 $18.00
Heavy production 50.0M tokens 25.0M tokens $15.00 $180.00

Need to model your own usage? Use the LLM pricing calculator →

Gemini 2.0 Flash features

Alternatives to Gemini 2.0 Flash

Cheaper, same-tier alternatives

If Gemini 2.0 Flash is more than you need, these budget-tier models cost less per token:

Higher-priced same-tier models

If you've outgrown Gemini 2.0 Flash, consider these:

Need the cheapest Google option? Gemini 2.5 Flash starts at $ 0.30 input / $ 2.50 output per 1M tokens →

FAQ

How much does Gemini 2.0 Flash cost per 1M tokens?

Gemini 2.0 Flash costs $0.10 per 1M input tokens and $0.40 per 1M output tokens. Cached input drops to $0.01 per 1M tokens — a 90% discount on repeated context.

What is Gemini 2.0 Flash's context window?

Gemini 2.0 Flash supports a 1,000,000-token context window with a maximum output of 8,000 tokens per response.

What features does Gemini 2.0 Flash support?

Vision, Tool Use, Streaming, Audio.

How much does Gemini 2.0 Flash cost for medium production usage?

For ~5M input + 2.5M output tokens per month, Gemini 2.0 Flash costs approximately $1.50/mo (≈$ 18.00/yr).

What are cheaper alternatives to Gemini 2.0 Flash?

Same-tier alternatives that cost less: Mistral Small 3.1 , Llama 3.3 70B , DeepSeek-V3 .