Directory

AI model API pricing

Every current, legacy and deprecated text and multimodal model we track across 11 providers, with input, output and cached token rates in USD per 1M tokens. Prices checked 2026-07-02.

    • Context
    • Input$0.33 / 1M tokens
    • Output$2.75 / 1M tokens
    • Cached$0.0825 / 1M tokens
    • Context
    • Input$0.30 / 1M tokens
    • Output$2.80 / 1M tokens
    • Cached
  • Amazon Nova 2.0 Pro

    Amazon (Nova)

    • Context
    • Input$1.38 / 1M tokens
    • Output$11.00 / 1M tokens
    • Cached
  • Amazon Nova Lite

    Amazon (Nova)

    • Context
    • Input$0.06 / 1M tokens
    • Output$0.24 / 1M tokens
    • Cached$0.015 / 1M tokens
  • Amazon Nova Micro

    Amazon (Nova)

    • Context
    • Input$0.035 / 1M tokens
    • Output$0.14 / 1M tokens
    • Cached$0.0088 / 1M tokens
  • Amazon Nova Premier

    Amazon (Nova)

    • Context
    • Input$2.50 / 1M tokens
    • Output$12.50 / 1M tokens
    • Cached$0.625 / 1M tokens
  • Amazon Nova Pro

    Amazon (Nova)

    • Context
    • Input$0.80 / 1M tokens
    • Output$3.20 / 1M tokens
    • Cached$0.20 / 1M tokens
    • Context
    • Input$1.00 / 1M tokens
    • Output$4.00 / 1M tokens
    • Cached
    • Context
    • Input$0.50 / 1M tokens
    • Output$1.50 / 1M tokens
    • Cached
    • Context
    • Input$0.50 / 1M tokens
    • Output$1.50 / 1M tokens
    • Cached
    • Context400K
    • Input$5.00 / 1M tokens
    • Output$30.00 / 1M tokens
    • Cached$0.50 / 1M tokens
    • Context1M
    • Input$10.00 / 1M tokens
    • Output$50.00 / 1M tokens
    • Cached$1.00 / 1M tokens
    • Context200K
    • Input$1.00 / 1M tokens
    • Output$5.00 / 1M tokens
    • Cached$0.10 / 1M tokens
    • Context1M
    • Input$10.00 / 1M tokens
    • Output$50.00 / 1M tokens
    • Cached$1.00 / 1M tokens
    • Context1M
    • Input$5.00 / 1M tokens
    • Output$25.00 / 1M tokens
    • Cached$0.50 / 1M tokens
    • Context1M
    • Input$2.00 / 1M tokens
    • Output$10.00 / 1M tokens
    • Cached$0.20 / 1M tokens
  • Codestral

    Mistral AI

    • Context
    • Input$0.30 / 1M tokens
    • Output$0.90 / 1M tokens
    • Cached
  • Command A

    Cohere

    Unverified
    • Context256K
    • Input
    • Output
    • Cached
  • Unverified
    • Context256K
    • Input
    • Output
    • Cached
  • Unverified
    • Context8K
    • Input
    • Output
    • Cached
  • Unverified
    • Context128K
    • Input
    • Output
    • Cached
  • Unverified
    • Context128K
    • Input
    • Output
    • Cached
  • Unverified
    • Context128K
    • Input
    • Output
    • Cached
  • Unverified
    • Context128K
    • Input
    • Output
    • Cached
    • Context1M
    • Input$0.14 / 1M tokens
    • Output$0.28 / 1M tokens
    • Cached$0.0028 / 1M tokens
    • Context1M
    • Input$0.435 / 1M tokens
    • Output$0.87 / 1M tokens
    • Cached$0.0036 / 1M tokens
  • Devstral 2

    Mistral AI

    • Context
    • Input$0.40 / 1M tokens
    • Output$2.00 / 1M tokens
    • Cached
    • Context
    • Input$0.10 / 1M tokens
    • Output$0.30 / 1M tokens
    • Cached
    • Context
    • Input$1.25 / 1M tokens
    • Output$10.00 / 1M tokens
    • Cached
    • Context
    • Input$0.30 / 1M tokens
    • Output$2.50 / 1M tokens
    • Cached$0.03 / 1M tokens
    • Context
    • Input$0.10 / 1M tokens
    • Output$0.40 / 1M tokens
    • Cached$0.01 / 1M tokens
    • Context
    • Input$1.25 / 1M tokens
    • Output$10.00 / 1M tokens
    • Cached$0.125 / 1M tokens
    • Context
    • Input$0.50 / 1M tokens
    • Output$3.00 / 1M tokens
    • Cached$0.05 / 1M tokens
    • Context
    • Input$0.25 / 1M tokens
    • Output$1.50 / 1M tokens
    • Cached$0.025 / 1M tokens
    • Context
    • Input$2.00 / 1M tokens
    • Output$12.00 / 1M tokens
    • Cached$0.20 / 1M tokens
    • Context1M
    • Input$1.50 / 1M tokens
    • Output$9.00 / 1M tokens
    • Cached$0.15 / 1M tokens
    • Context400K
    • Input$1.75 / 1M tokens
    • Output$14.00 / 1M tokens
    • Cached$0.175 / 1M tokens
  • GPT-5.4

    OpenAI

    • Context1.1M
    • Input$2.50 / 1M tokens
    • Output$15.00 / 1M tokens
    • Cached$0.25 / 1M tokens
    • Context400K
    • Input$0.75 / 1M tokens
    • Output$4.50 / 1M tokens
    • Cached$0.075 / 1M tokens
    • Context400K
    • Input$0.20 / 1M tokens
    • Output$1.25 / 1M tokens
    • Cached$0.02 / 1M tokens
  • GPT-5.5

    OpenAI

    • Context1.1M
    • Input$5.00 / 1M tokens
    • Output$30.00 / 1M tokens
    • Cached$0.50 / 1M tokens
    • Context1.1M
    • Input$30.00 / 1M tokens
    • Output$180.00 / 1M tokens
    • Cached
    • Context1M
    • Input$1.25 / 1M tokens
    • Output$2.50 / 1M tokens
    • Cached$0.20 / 1M tokens
    • Context1M
    • Input$1.25 / 1M tokens
    • Output$2.50 / 1M tokens
    • Cached$0.20 / 1M tokens
    • Context1M
    • Input$1.25 / 1M tokens
    • Output$2.50 / 1M tokens
    • Cached$0.20 / 1M tokens
    • Context1M
    • Input$1.25 / 1M tokens
    • Output$2.50 / 1M tokens
    • Cached$0.20 / 1M tokens
    • Context256K
    • Input$1.00 / 1M tokens
    • Output$2.00 / 1M tokens
    • Cached
  • Leanstral

    Mistral AI

    • Context
    • Input$0.00 / 1M tokens
    • Output$0.00 / 1M tokens
    • Cached
    • Context1M
    • Input$0.15 / 1M tokens
    • Output$0.60 / 1M tokens
    • Cached
    • Context10M
    • Input$0.10 / 1M tokens
    • Output$0.30 / 1M tokens
    • Cached
    • Context
    • Input$2.00 / 1M tokens
    • Output$5.00 / 1M tokens
    • Cached
  • Magistral Small

    Mistral AI

    • Context
    • Input$0.50 / 1M tokens
    • Output$1.50 / 1M tokens
    • Cached
    • Context
    • Input$0.20 / 1M tokens
    • Output$0.20 / 1M tokens
    • Cached
    • Context
    • Input$0.10 / 1M tokens
    • Output$0.10 / 1M tokens
    • Cached
    • Context
    • Input$0.15 / 1M tokens
    • Output$0.15 / 1M tokens
    • Cached
  • Mistral Large 3

    Mistral AI

    • Context
    • Input$0.50 / 1M tokens
    • Output$1.50 / 1M tokens
    • Cached
    • Context
    • Input$1.50 / 1M tokens
    • Output$7.50 / 1M tokens
    • Cached
    • Context
    • Input$0.10 / 1M tokens
    • Output
    • Cached
  • Mistral OCR 4

    Mistral AI

    • Context
    • Input
    • Output
    • Cached
  • Mistral Small 4

    Mistral AI

    • Context
    • Input$0.15 / 1M tokens
    • Output$0.60 / 1M tokens
    • Cached
  • Qwen-Flash

    Alibaba Cloud (Qwen)

    • Context
    • Input$0.05 / 1M tokens
    • Output$0.40 / 1M tokens
    • Cached
  • Qwen-Long

    Alibaba Cloud (Qwen)

    • Context
    • Input$0.072 / 1M tokens
    • Output$0.287 / 1M tokens
    • Cached
  • Qwen-MT-Flash

    Alibaba Cloud (Qwen)

    • Context
    • Input$0.16 / 1M tokens
    • Output$0.49 / 1M tokens
    • Cached
  • Qwen-Plus

    Alibaba Cloud (Qwen)

    • Context
    • Input$0.115 / 1M tokens
    • Output$1.15 / 1M tokens
    • Cached
  • Qwen3-Max

    Alibaba Cloud (Qwen)

    • Context
    • Input$1.20 / 1M tokens
    • Output$6.00 / 1M tokens
    • Cached
  • Qwen3-Omni-Flash

    Alibaba Cloud (Qwen)

    Unverified
    • Context
    • Input
    • Output
    • Cached
  • Qwen3-VL-Plus

    Alibaba Cloud (Qwen)

    • Context
    • Input$0.20 / 1M tokens
    • Output$1.60 / 1M tokens
    • Cached
  • Qwen3.5-Flash

    Alibaba Cloud (Qwen)

    • Context
    • Input$0.10 / 1M tokens
    • Output$0.40 / 1M tokens
    • Cached
  • Qwen3.6-Flash

    Alibaba Cloud (Qwen)

    • Context
    • Input$0.25 / 1M tokens
    • Output$1.50 / 1M tokens
    • Cached
  • Qwen3.7-Max

    Alibaba Cloud (Qwen)

    • Context
    • Input$2.50 / 1M tokens
    • Output$7.50 / 1M tokens
    • Cached
  • Qwen3.7-Plus

    Alibaba Cloud (Qwen)

    • Context
    • Input$0.40 / 1M tokens
    • Output$1.60 / 1M tokens
    • Cached
  • QwQ-Plus

    Alibaba Cloud (Qwen)

    • Context
    • Input$0.80 / 1M tokens
    • Output$2.40 / 1M tokens
    • Cached
  • Unverified
    • Context4K
    • Input
    • Output
    • Cached
  • Unverified
    • Context4K
    • Input
    • Output
    • Cached
  • Unverified
    • Context4K
    • Input
    • Output
    • Cached
  • Unverified
    • Context32K
    • Input
    • Output
    • Cached
  • Unverified
    • Context32K
    • Input
    • Output
    • Cached
  • rerank-2.5

    Voyage AI

    • Context
    • Input$0.05 / 1M tokens
    • Output
    • Cached
    • Context
    • Input$0.02 / 1M tokens
    • Output
    • Cached
  • Legacy
    • Context200K
    • Input$5.00 / 1M tokens
    • Output$25.00 / 1M tokens
    • Cached$0.50 / 1M tokens
  • Legacy
    • Context1M
    • Input$5.00 / 1M tokens
    • Output$25.00 / 1M tokens
    • Cached$0.50 / 1M tokens
  • Legacy
    • Context1M
    • Input$5.00 / 1M tokens
    • Output$25.00 / 1M tokens
    • Cached$0.50 / 1M tokens
  • Legacy
    • Context200K
    • Input$3.00 / 1M tokens
    • Output$15.00 / 1M tokens
    • Cached$0.30 / 1M tokens
  • Legacy
    • Context1M
    • Input$3.00 / 1M tokens
    • Output$15.00 / 1M tokens
    • Cached$0.30 / 1M tokens
    • Context128K
    • Input$2.50 / 1M tokens
    • Output$10.00 / 1M tokens
    • Cached
  • GPT-4.1

    OpenAI

    Legacy
    • Context1M
    • Input$2.00 / 1M tokens
    • Output$8.00 / 1M tokens
    • Cached$0.50 / 1M tokens
  • Legacy
    • Context1M
    • Input$0.40 / 1M tokens
    • Output$1.60 / 1M tokens
    • Cached$0.10 / 1M tokens
  • Legacy
    • Context1M
    • Input$0.10 / 1M tokens
    • Output$0.40 / 1M tokens
    • Cached$0.025 / 1M tokens
  • GPT-4o

    OpenAI

    Legacy
    • Context128K
    • Input$2.50 / 1M tokens
    • Output$10.00 / 1M tokens
    • Cached$1.25 / 1M tokens
  • Legacy
    • Context128K
    • Input$0.15 / 1M tokens
    • Output$0.60 / 1M tokens
    • Cached$0.075 / 1M tokens
  • Legacy
    • Context1.1M
    • Input$30.00 / 1M tokens
    • Output$180.00 / 1M tokens
    • Cached
    • Context131K
    • Input$0.10 / 1M tokens
    • Output$0.32 / 1M tokens
    • Cached
  • Mistral NeMo

    Mistral AI

    Legacy
    • Context
    • Input$0.15 / 1M tokens
    • Output$0.15 / 1M tokens
    • Cached
  • Mixtral 8x22B

    Mistral AI

    Legacy
    • Context
    • Input$2.00 / 1M tokens
    • Output$6.00 / 1M tokens
    • Cached
  • Mixtral 8x7B

    Mistral AI

    Legacy
    • Context
    • Input$0.70 / 1M tokens
    • Output$0.70 / 1M tokens
    • Cached
  • o3

    OpenAI

    Legacy
    • Context200K
    • Input$2.00 / 1M tokens
    • Output$8.00 / 1M tokens
    • Cached$0.50 / 1M tokens
  • Legacy
    • Context
    • Input
    • Output
    • Cached
  • o3-mini

    OpenAI

    Legacy
    • Context200K
    • Input$1.10 / 1M tokens
    • Output$4.40 / 1M tokens
    • Cached$0.55 / 1M tokens
  • o4-mini

    OpenAI

    Legacy
    • Context200K
    • Input$1.10 / 1M tokens
    • Output$4.40 / 1M tokens
    • Cached$0.275 / 1M tokens
    • Context
    • Input
    • Output
    • Cached
  • Qwen-Turbo

    Alibaba Cloud (Qwen)

    Legacy
    • Context
    • Input$0.05 / 1M tokens
    • Output$0.20 / 1M tokens
    • Cached
  • rerank-2

    Voyage AI

    Legacy
    • Context
    • Input$0.05 / 1M tokens
    • Output
    • Cached
  • rerank-2-lite

    Voyage AI

    Legacy
    • Context
    • Input$0.02 / 1M tokens
    • Output
    • Cached
  • DeprecatedUnverified
    • Context
    • Input$0.80 / 1M tokens
    • Output$4.00 / 1M tokens
    • Cached$0.08 / 1M tokens
  • Claude Opus 4

    Anthropic

    DeprecatedUnverified
    • Context
    • Input$15.00 / 1M tokens
    • Output$75.00 / 1M tokens
    • Cached$1.50 / 1M tokens
  • Deprecated
    • Context200K
    • Input$15.00 / 1M tokens
    • Output$75.00 / 1M tokens
    • Cached$1.50 / 1M tokens
  • DeprecatedUnverified
    • Context
    • Input$3.00 / 1M tokens
    • Output$15.00 / 1M tokens
    • Cached$0.30 / 1M tokens
  • Command

    Cohere

    Deprecated
    • Context
    • Input$1.00 / 1M tokens
    • Output$2.00 / 1M tokens
    • Cached
  • Deprecated
    • Context
    • Input$0.30 / 1M tokens
    • Output$0.60 / 1M tokens
    • Cached
  • Deprecated
    • Context
    • Input$0.50 / 1M tokens
    • Output$1.50 / 1M tokens
    • Cached
  • Deprecated
    • Context
    • Input$3.00 / 1M tokens
    • Output$15.00 / 1M tokens
    • Cached
  • Deprecated
    • Context
    • Input$0.10 / 1M tokens
    • Output$0.40 / 1M tokens
    • Cached
  • Deprecated
    • Context
    • Input$0.075 / 1M tokens
    • Output$0.30 / 1M tokens
    • Cached
  • o1

    OpenAI

    Deprecated
    • Context200K
    • Input$15.00 / 1M tokens
    • Output$60.00 / 1M tokens
    • Cached$7.50 / 1M tokens
  • rerank-1

    Voyage AI

    Deprecated
    • Context
    • Input$0.05 / 1M tokens
    • Output
    • Cached
  • rerank-lite-1

    Voyage AI

    Deprecated
    • Context
    • Input$0.02 / 1M tokens
    • Output
    • Cached

Embedding models

Priced per input token only — there is no output leg for embeddings.

Image & audio models (15)

Usually billed per image, per second or with separate speech/text rates rather than per token — see each model’s page for its documented unit price.