Gemini 3.1 Pro Preview API Pricing
Tiered pricing: rates shown are for prompts ≤200K tokens. Above 200K tokens: input $4.00/M, output $18.00/M, cached input $0.40/M. Context caching storage billed separately at $4.50/hour.
Verified against Google’s pricing page on 2026-07-02.
View source ↗Facts
Specs
- FamilyGemini
- API IDgemini-3.1-pro-preview
- ModalityMultimodal
- Context window—
- Max output—
- Release dateNot published
Pricing (per 1M tokens)
- Input$2.00 / 1M tokens
- Output$12.00 / 1M tokens
- Cached input$0.20 / 1M tokens
- Cache write—
- Batch discount50%
Capabilities
vision, tool-use, multimodal
Notes
Tiered pricing: rates shown are for prompts ≤200K tokens. Above 200K tokens: input $4.00/M, output $18.00/M, cached input $0.40/M. Context caching storage billed separately at $4.50/hour.
Worked cost examples
Estimated monthly cost on three common usage shapes
| Workload | Est. monthly cost |
|---|---|
Chatbot 800 input / 200 output tokens per request, 100,000 requests/month | $400.00 |
RAG pipeline 6,000 input / 700 output tokens per request, 30,000 requests/month | $612.00 |
Batch summarization 20,000 input / 1,500 output tokens per request, 5,000 requests/month | $290.00 |
Price history
Changes we have documented since we began tracking
No price changes recorded since we began tracking.
Gemini 3.1 Pro Preview pricing FAQ
- How much does Gemini 3.1 Pro Preview cost per 1M tokens?
- Gemini 3.1 Pro Preview costs $2.00 per 1M input tokens and $12.00 per 1M output tokens.
- How much does Gemini 3.1 Pro Preview cost per request?
- A typical request of 800 input and 200 output tokens costs about $0.004 on Gemini 3.1 Pro Preview.
- Is Gemini 3.1 Pro Preview cheaper than Gemini 2.0 Flash?
- Gemini 3.1 Pro Preview is more expensive than Gemini 2.0 Flash on input tokens, by about 1900%.
- Does Gemini 3.1 Pro Preview have cache pricing?
- Yes — cached input reads cost $0.20 per 1M tokens.
Keep exploring
Provider
All Google pricing →Related models