DeepSeek V4 Flash pricing
DeepSeek · list prices as of 2026-09-29
| Input | $0.30 per 1M tokens |
|---|---|
| Output | $1.20 per 1M tokens |
| Cached input (read) | $0.006 per 1M tokens |
| Context window | 1M tokens |
| Longest answer | 393K tokens |
What a call costs
| Typical call | Input tokens | Output tokens | Per call | Per month at 1,000 a day |
|---|---|---|---|---|
| Chat reply | 1,000 | 300 | $0.00066 | $19.80 |
| Document summary | 6,000 | 400 | $0.00228 | $68.40 |
| Code change | 3,000 | 1,200 | $0.00234 | $70.20 |
No caching or batch discount. Your own prompt will differ: count it exactly below.
Price your own prompt on DeepSeek V4 FlashSimilarly priced models
- Gemini 3.1 Flash-Lite (Google): 1.1× more
- Gemini 3.1 Flash-Lite (preview) (Google): 1.1× more
- DeepSeek V3 (DeepSeek): 1.1× cheaper
- GPT-5.4 nano (OpenAI): 1.1× cheaper
- Llama 4 Maverick (Meta): 1.3× cheaper
Questions
How much does DeepSeek V4 Flash cost?
DeepSeek V4 Flash costs $0.30 per million input tokens and $1.20 per million output tokens, as of 2026-09-29.
How much is one DeepSeek V4 Flash call?
A call with 1,000 input tokens and a 300-token answer costs about $0.00066, or $19.80 a month at 1,000 calls a day.
How big is the DeepSeek V4 Flash context window?
1,000,000 tokens.