Google

Gemini 3.1 Pro Preview API Pricing

Tiered pricing: rates shown are for prompts ≤200K tokens. Above 200K tokens: input $4.00/M, output $18.00/M, cached input $0.40/M. Context caching storage billed separately at $4.50/hour.

Verified against Google’s pricing page on 2026-07-02.

View source ↗

Facts

Specs

  • FamilyGemini
  • API IDgemini-3.1-pro-preview
  • ModalityMultimodal
  • Context window
  • Max output
  • Release dateNot published

Pricing (per 1M tokens)

  • Input$2.00 / 1M tokens
  • Output$12.00 / 1M tokens
  • Cached input$0.20 / 1M tokens
  • Cache write
  • Batch discount50%

Capabilities

vision, tool-use, multimodal

Notes

Tiered pricing: rates shown are for prompts ≤200K tokens. Above 200K tokens: input $4.00/M, output $18.00/M, cached input $0.40/M. Context caching storage billed separately at $4.50/hour.

Worked cost examples

Estimated monthly cost on three common usage shapes

WorkloadEst. monthly cost

Chatbot

800 input / 200 output tokens per request, 100,000 requests/month

$400.00

RAG pipeline

6,000 input / 700 output tokens per request, 30,000 requests/month

$612.00

Batch summarization

20,000 input / 1,500 output tokens per request, 5,000 requests/month

$290.00

Price history

Changes we have documented since we began tracking

No price changes recorded since we began tracking.

Gemini 3.1 Pro Preview pricing FAQ

How much does Gemini 3.1 Pro Preview cost per 1M tokens?
Gemini 3.1 Pro Preview costs $2.00 per 1M input tokens and $12.00 per 1M output tokens.
How much does Gemini 3.1 Pro Preview cost per request?
A typical request of 800 input and 200 output tokens costs about $0.004 on Gemini 3.1 Pro Preview.
Is Gemini 3.1 Pro Preview cheaper than Gemini 2.0 Flash?
Gemini 3.1 Pro Preview is more expensive than Gemini 2.0 Flash on input tokens, by about 1900%.
Does Gemini 3.1 Pro Preview have cache pricing?
Yes — cached input reads cost $0.20 per 1M tokens.