LLMCost
deepseek

DeepSeek: DeepSeek V4 Flash 0731 API Pricing

DeepSeek: DeepSeek V4 Flash 0731 costs $0.080 per 1M input tokens and $0.180 per 1M output tokens (cache reads: $0.016). Context window: 1.3M. Prices verified 2026-08-23 · source

Input / 1M tokens
$0.080
Output / 1M tokens
$0.180
Context window
1.3M
Max output
384K

What would DeepSeek: DeepSeek V4 Flash 0731 cost you per month?

WorkloadRequests/moAvg tokens (in/out)Est. monthly cost
Light chatbot10,000800 / 400$1.36
Coding agent5,0008,000 / 2,000$5.00
RAG pipeline30,0004,000 / 500$12.30
Doc summarizer2,00020,000 / 1,000$3.56

Estimates use list prices without prompt caching. Caching can cut input cost by up to 90% on repeated context.

About DeepSeek: DeepSeek V4 Flash 0731

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....

Compare DeepSeek: DeepSeek V4 Flash 0731 with