Gemini 2.0 Flash API pricing (August 2026)
Shut down by Google on 1 Jun 2026 — shown at its final price for historical comparison only.
What it costs at real volumes
A chat-shaped workload: 800 input tokens and 300 output tokens per request, no caching, no batch.
Specifications
- Provider
- Google Gemini
- Context window
- 1.0M tokens
- Max output
- 8.2K tokens
- Tier
- small
- Status
- DEPRECATED
- Blended $/1M (3:1)
- $0.175
- Cache break-even hit rate
- 0% — caching always saves
- Prices verified
- 4 Aug 2026
Source: Google Gemini pricing. Finitizer is not affiliated with Google Gemini.
Similar models
Other gemini models
- Gemini 2.5 Pro+$3.26
- Gemini 2.5 Flash+$0.675
- Gemini 2.5 Flash-Lite+$0
Closest prices elsewhere
- GPT-4.1 nano+$0
- DeepSeek V4 Flash+$0
- GPT-5 nano$-0.038
- Amazon Nova Lite$-0.070
Frequently asked questions
How much does Gemini 2.0 Flash cost per 1M tokens?
Gemini 2.0 Flash costs $0.100 per million input tokens and $0.400 per million output tokens at Google Gemini list prices. Output is 4.0x the input rate, so the input:output ratio of your workload drives the real cost more than the headline price.
Does Gemini 2.0 Flash support prompt caching?
Yes. Cached input is billed at $0.025 per million tokens, 4x cheaper than uncached input. There is no cache-write premium, so caching is never a net loss.
What is the batch discount for Gemini 2.0 Flash?
Requests submitted through the asynchronous batch endpoint are billed at 50% off both input and output rates, bringing input to $0.050 and output to $0.200 per million tokens.
What is Gemini 2.0 Flash's context window?
1.0M tokens of context with up to 8.2K output tokens per request. A larger context window raises the ceiling on what you can send, not the price per token — but filling it on every request is one of the most common causes of an unexpected bill.
Running Gemini 2.0 Flash in production?
Finitizer TokenOps attributes real token spend to teams, features, and prompts — so you find out which workload moved the bill before the invoice does.