DeepSeek
DeepSeek V4 Pro API Pricing
inputPerMTok is the cache-miss (standard) input rate; cachedInputPerMTok is the cache-hit rate. Concurrency limit 500.
Verified against DeepSeek’s pricing page on 2026-07-02.
View source ↗Facts
Specs
- FamilyDeepSeek V4
- API IDdeepseek-v4-pro
- ModalityText
- Context window1M tokens
- Max output384K tokens
- Release dateNot published
Pricing (per 1M tokens)
- Input$0.435 / 1M tokens
- Output$0.87 / 1M tokens
- Cached input$0.0036 / 1M tokens
- Cache write—
- Batch discount—
Capabilities
tool-use, reasoning, json-mode
Notes
inputPerMTok is the cache-miss (standard) input rate; cachedInputPerMTok is the cache-hit rate. Concurrency limit 500.
Worked cost examples
Estimated monthly cost on three common usage shapes
| Workload | Est. monthly cost |
|---|---|
Chatbot 800 input / 200 output tokens per request, 100,000 requests/month | $52.20 |
RAG pipeline 6,000 input / 700 output tokens per request, 30,000 requests/month | $96.57 |
Batch summarization 20,000 input / 1,500 output tokens per request, 5,000 requests/month | $50.03 |
Price history
Changes we have documented since we began tracking
No price changes recorded since we began tracking.
DeepSeek V4 Pro pricing FAQ
- How much does DeepSeek V4 Pro cost per 1M tokens?
- DeepSeek V4 Pro costs $0.435 per 1M input tokens and $0.87 per 1M output tokens.
- How much does DeepSeek V4 Pro cost per request?
- A typical request of 800 input and 200 output tokens costs about $0.000522 on DeepSeek V4 Pro.
- Is DeepSeek V4 Pro cheaper than DeepSeek V4 Flash?
- DeepSeek V4 Pro is more expensive than DeepSeek V4 Flash on input tokens, by about 211%.
- Does DeepSeek V4 Pro have cache pricing?
- Yes — cached input reads cost $0.0036 per 1M tokens.