Google

Gemini 3.1 Flash-Lite API Pricing

Rate shown is for text/image/video input; audio input is priced at $0.50/M. Batch input $0.125/M (audio $0.25/M), batch output $0.75/M.

Verified against Google’s pricing page on 2026-07-02.

View source ↗

Facts

Specs

  • FamilyGemini
  • API IDgemini-3.1-flash-lite
  • ModalityMultimodal
  • Context window
  • Max output
  • Release dateNot published

Pricing (per 1M tokens)

  • Input$0.25 / 1M tokens
  • Output$1.50 / 1M tokens
  • Cached input$0.025 / 1M tokens
  • Cache write
  • Batch discount50%

Capabilities

vision, tool-use, multimodal

Notes

Rate shown is for text/image/video input; audio input is priced at $0.50/M. Batch input $0.125/M (audio $0.25/M), batch output $0.75/M.

Worked cost examples

Estimated monthly cost on three common usage shapes

WorkloadEst. monthly cost

Chatbot

800 input / 200 output tokens per request, 100,000 requests/month

$50.00

RAG pipeline

6,000 input / 700 output tokens per request, 30,000 requests/month

$76.50

Batch summarization

20,000 input / 1,500 output tokens per request, 5,000 requests/month

$36.25

Price history

Changes we have documented since we began tracking

No price changes recorded since we began tracking.

Gemini 3.1 Flash-Lite pricing FAQ

How much does Gemini 3.1 Flash-Lite cost per 1M tokens?
Gemini 3.1 Flash-Lite costs $0.25 per 1M input tokens and $1.50 per 1M output tokens.
How much does Gemini 3.1 Flash-Lite cost per request?
A typical request of 800 input and 200 output tokens costs about $0.0005 on Gemini 3.1 Flash-Lite.
Is Gemini 3.1 Flash-Lite cheaper than Gemini 2.0 Flash?
Gemini 3.1 Flash-Lite is more expensive than Gemini 2.0 Flash on input tokens, by about 150%.
Does Gemini 3.1 Flash-Lite have cache pricing?
Yes — cached input reads cost $0.025 per 1M tokens.