DeepSeek
DeepSeek V4 Flash API Pricing
inputPerMTok is the cache-miss (standard) input rate; cachedInputPerMTok is the cache-hit rate. Non-thinking mode is the legacy alias `deepseek-chat`; thinking mode is the legacy alias `deepseek-reasoner` (both aliases scheduled for deprecation 2026-07-24, after which they route here). Concurrency limit 2,500.
Verified against DeepSeek’s pricing page on 2026-07-02.
View source ↗Facts
Specs
- FamilyDeepSeek V4
- API IDdeepseek-v4-flash
- ModalityText
- Context window1M tokens
- Max output384K tokens
- Release dateNot published
Pricing (per 1M tokens)
- Input$0.14 / 1M tokens
- Output$0.28 / 1M tokens
- Cached input$0.0028 / 1M tokens
- Cache write—
- Batch discount—
Capabilities
tool-use, reasoning, json-mode
Notes
inputPerMTok is the cache-miss (standard) input rate; cachedInputPerMTok is the cache-hit rate. Non-thinking mode is the legacy alias `deepseek-chat`; thinking mode is the legacy alias `deepseek-reasoner` (both aliases scheduled for deprecation 2026-07-24, after which they route here). Concurrency limit 2,500.
Worked cost examples
Estimated monthly cost on three common usage shapes
| Workload | Est. monthly cost |
|---|---|
Chatbot 800 input / 200 output tokens per request, 100,000 requests/month | $16.80 |
RAG pipeline 6,000 input / 700 output tokens per request, 30,000 requests/month | $31.08 |
Batch summarization 20,000 input / 1,500 output tokens per request, 5,000 requests/month | $16.10 |
Price history
Changes we have documented since we began tracking
No price changes recorded since we began tracking.
DeepSeek V4 Flash pricing FAQ
- How much does DeepSeek V4 Flash cost per 1M tokens?
- DeepSeek V4 Flash costs $0.14 per 1M input tokens and $0.28 per 1M output tokens.
- How much does DeepSeek V4 Flash cost per request?
- A typical request of 800 input and 200 output tokens costs about $0.000168 on DeepSeek V4 Flash.
- Is DeepSeek V4 Flash cheaper than DeepSeek V4 Pro?
- DeepSeek V4 Flash is cheaper than DeepSeek V4 Pro on input tokens, by about 68%.
- Does DeepSeek V4 Flash have cache pricing?
- Yes — cached input reads cost $0.0028 per 1M tokens.