DeepSeek

DeepSeek V4 Pro API Pricing

inputPerMTok is the cache-miss (standard) input rate; cachedInputPerMTok is the cache-hit rate. Concurrency limit 500.

Verified against DeepSeek’s pricing page on 2026-07-02.

View source ↗

Facts

Specs

  • FamilyDeepSeek V4
  • API IDdeepseek-v4-pro
  • ModalityText
  • Context window1M tokens
  • Max output384K tokens
  • Release dateNot published

Pricing (per 1M tokens)

  • Input$0.435 / 1M tokens
  • Output$0.87 / 1M tokens
  • Cached input$0.0036 / 1M tokens
  • Cache write
  • Batch discount

Capabilities

tool-use, reasoning, json-mode

Notes

inputPerMTok is the cache-miss (standard) input rate; cachedInputPerMTok is the cache-hit rate. Concurrency limit 500.

Worked cost examples

Estimated monthly cost on three common usage shapes

WorkloadEst. monthly cost

Chatbot

800 input / 200 output tokens per request, 100,000 requests/month

$52.20

RAG pipeline

6,000 input / 700 output tokens per request, 30,000 requests/month

$96.57

Batch summarization

20,000 input / 1,500 output tokens per request, 5,000 requests/month

$50.03

Price history

Changes we have documented since we began tracking

No price changes recorded since we began tracking.

DeepSeek V4 Pro pricing FAQ

How much does DeepSeek V4 Pro cost per 1M tokens?
DeepSeek V4 Pro costs $0.435 per 1M input tokens and $0.87 per 1M output tokens.
How much does DeepSeek V4 Pro cost per request?
A typical request of 800 input and 200 output tokens costs about $0.000522 on DeepSeek V4 Pro.
Is DeepSeek V4 Pro cheaper than DeepSeek V4 Flash?
DeepSeek V4 Pro is more expensive than DeepSeek V4 Flash on input tokens, by about 211%.
Does DeepSeek V4 Pro have cache pricing?
Yes — cached input reads cost $0.0036 per 1M tokens.