Gemini 2.0 Flash API Pricing
Google's pricing page lists this model as shut down June 1, 2026 (already past as of verification date). Audio input was priced at $0.70/M.
Verified against Google’s pricing page on 2026-07-02.
View source ↗Facts
Specs
- FamilyGemini
- API IDgemini-2.0-flash
- ModalityMultimodal
- Context window—
- Max output—
- Release dateNot published
Pricing (per 1M tokens)
- Input$0.10 / 1M tokens
- Output$0.40 / 1M tokens
- Cached input—
- Cache write—
- Batch discount50%
Capabilities
vision, tool-use, multimodal
Notes
Google's pricing page lists this model as shut down June 1, 2026 (already past as of verification date). Audio input was priced at $0.70/M.
Worked cost examples
Estimated monthly cost on three common usage shapes
| Workload | Est. monthly cost |
|---|---|
Chatbot 800 input / 200 output tokens per request, 100,000 requests/month | $16.00 |
RAG pipeline 6,000 input / 700 output tokens per request, 30,000 requests/month | $26.40 |
Batch summarization 20,000 input / 1,500 output tokens per request, 5,000 requests/month | $13.00 |
Price history
Changes we have documented since we began tracking
No price changes recorded since we began tracking.
Gemini 2.0 Flash pricing FAQ
- How much does Gemini 2.0 Flash cost per 1M tokens?
- Gemini 2.0 Flash costs $0.10 per 1M input tokens and $0.40 per 1M output tokens.
- How much does Gemini 2.0 Flash cost per request?
- A typical request of 800 input and 200 output tokens costs about $0.00016 on Gemini 2.0 Flash.
- Is Gemini 2.0 Flash cheaper than Gemini 2.0 Flash-Lite?
- Gemini 2.0 Flash is more expensive than Gemini 2.0 Flash-Lite on input tokens, by about 33%.
- Does Gemini 2.0 Flash have cache pricing?
- No cache discount is documented for Gemini 2.0 Flash on the provider's pricing page.
Keep exploring
Provider
All Google pricing →Related models