Directory
AI model API pricing
Every current, legacy and deprecated text and multimodal model we track across 11 providers, with input, output and cached token rates in USD per 1M tokens. Prices checked 2026-07-02.
All providersAlibaba Cloud (Qwen)Amazon (Nova)AnthropicCohereDeepSeekGoogleMetaMistral AIOpenAIVoyage AIxAI
| Model | Provider | Context | Input / 1M | Output / 1M | Cached / 1M |
|---|---|---|---|---|---|
| Gemini 2.5 Computer Use Preview | — | $1.25 | $10.00 | — | |
| Gemini 2.5 Flash | — | $0.30 | $2.50 | $0.03 | |
| Gemini 2.5 Flash-Lite | — | $0.10 | $0.40 | $0.01 | |
| Gemini 2.5 Pro | — | $1.25 | $10.00 | $0.125 | |
| Gemini 3 Flash Preview | — | $0.50 | $3.00 | $0.05 | |
| Gemini 3.1 Flash-Lite | — | $0.25 | $1.50 | $0.025 | |
| Gemini 3.1 Pro Preview | — | $2.00 | $12.00 | $0.20 | |
| Gemini 3.5 Flash | 1M | $1.50 | $9.00 | $0.15 | |
| Gemini 2.0 FlashDeprecated | — | $0.10 | $0.40 | — | |
| Gemini 2.0 Flash-LiteDeprecated | — | $0.075 | $0.30 | — |
- Context—
- Input$1.25 / 1M tokens
- Output$10.00 / 1M tokens
- Cached—
- Gemini 2.5 Flash
Google
- Context—
- Input$0.30 / 1M tokens
- Output$2.50 / 1M tokens
- Cached$0.03 / 1M tokens
- Gemini 2.5 Flash-Lite
Google
- Context—
- Input$0.10 / 1M tokens
- Output$0.40 / 1M tokens
- Cached$0.01 / 1M tokens
- Gemini 2.5 Pro
Google
- Context—
- Input$1.25 / 1M tokens
- Output$10.00 / 1M tokens
- Cached$0.125 / 1M tokens
- Gemini 3 Flash Preview
Google
- Context—
- Input$0.50 / 1M tokens
- Output$3.00 / 1M tokens
- Cached$0.05 / 1M tokens
- Gemini 3.1 Flash-Lite
Google
- Context—
- Input$0.25 / 1M tokens
- Output$1.50 / 1M tokens
- Cached$0.025 / 1M tokens
- Gemini 3.1 Pro Preview
Google
- Context—
- Input$2.00 / 1M tokens
- Output$12.00 / 1M tokens
- Cached$0.20 / 1M tokens
- Gemini 3.5 Flash
Google
- Context1M
- Input$1.50 / 1M tokens
- Output$9.00 / 1M tokens
- Cached$0.15 / 1M tokens
- Gemini 2.0 FlashDeprecated
Google
- Context—
- Input$0.10 / 1M tokens
- Output$0.40 / 1M tokens
- Cached—
- Gemini 2.0 Flash-LiteDeprecated
Google
- Context—
- Input$0.075 / 1M tokens
- Output$0.30 / 1M tokens
- Cached—
Embedding models
Priced per input token only — there is no output leg for embeddings.
| Model | Provider | Context | Input / 1M | Output / 1M | Cached / 1M |
|---|---|---|---|---|---|
| Gemini Embedding 001Legacy | — | $0.15 | — | — | |
| Gemini Embedding 2 | — | $0.20 | — | — |
- Gemini Embedding 001Legacy
Google
- Context—
- Input$0.15 / 1M tokens
- Output—
- Cached—
- Gemini Embedding 2
Google
- Context—
- Input$0.20 / 1M tokens
- Output—
- Cached—