deepseek
DeepSeek: DeepSeek V4 Flash 0731 API Pricing
DeepSeek: DeepSeek V4 Flash 0731 costs $0.080 per 1M input tokens and $0.180 per 1M output tokens (cache reads: $0.016). Context window: 1.3M. Prices verified 2026-08-23 · source
Input / 1M tokens
$0.080
Output / 1M tokens
$0.180
Context window
1.3M
Max output
384K
What would DeepSeek: DeepSeek V4 Flash 0731 cost you per month?
| Workload | Requests/mo | Avg tokens (in/out) | Est. monthly cost |
|---|---|---|---|
| Light chatbot | 10,000 | 800 / 400 | $1.36 |
| Coding agent | 5,000 | 8,000 / 2,000 | $5.00 |
| RAG pipeline | 30,000 | 4,000 / 500 | $12.30 |
| Doc summarizer | 2,000 | 20,000 / 1,000 | $3.56 |
Estimates use list prices without prompt caching. Caching can cut input cost by up to 90% on repeated context.
About DeepSeek: DeepSeek V4 Flash 0731
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....