Gemini 2.0 Flash-Lite API Pricing
Google's pricing page lists this model as shut down June 1, 2026 (already past as of verification date).
Verified against Google’s pricing page on 2026-07-02.
View source ↗Facts
Specs
- FamilyGemini
- API IDgemini-2.0-flash-lite
- ModalityMultimodal
- Context window—
- Max output—
- Release dateNot published
Pricing (per 1M tokens)
- Input$0.075 / 1M tokens
- Output$0.30 / 1M tokens
- Cached input—
- Cache write—
- Batch discount50%
Capabilities
vision, tool-use, multimodal
Notes
Google's pricing page lists this model as shut down June 1, 2026 (already past as of verification date).
Worked cost examples
Estimated monthly cost on three common usage shapes
| Workload | Est. monthly cost |
|---|---|
Chatbot 800 input / 200 output tokens per request, 100,000 requests/month | $12.00 |
RAG pipeline 6,000 input / 700 output tokens per request, 30,000 requests/month | $19.80 |
Batch summarization 20,000 input / 1,500 output tokens per request, 5,000 requests/month | $9.75 |
Price history
Changes we have documented since we began tracking
No price changes recorded since we began tracking.
Gemini 2.0 Flash-Lite pricing FAQ
- How much does Gemini 2.0 Flash-Lite cost per 1M tokens?
- Gemini 2.0 Flash-Lite costs $0.075 per 1M input tokens and $0.30 per 1M output tokens.
- How much does Gemini 2.0 Flash-Lite cost per request?
- A typical request of 800 input and 200 output tokens costs about $0.00012 on Gemini 2.0 Flash-Lite.
- Is Gemini 2.0 Flash-Lite cheaper than Gemini 2.0 Flash?
- Gemini 2.0 Flash-Lite is cheaper than Gemini 2.0 Flash on input tokens, by about 25%.
- Does Gemini 2.0 Flash-Lite have cache pricing?
- No cache discount is documented for Gemini 2.0 Flash-Lite on the provider's pricing page.
Keep exploring
Provider
All Google pricing →Related models