← All models
Flash — fastest and cheapest, lighter reasoningOpen Weights
DeepSeek V4 Flash 0731 (batch) Pricing
DeepSeek · deepseek-deepseek-v4-flash-0731-batch
Input / 1M tokens
$0.110
Output / 1M tokens
$0.330
Context window
1,048,576 tokens
≈ 1,398 pages of text
Max output
943,718 tokens
What it costs in practice
Typical request (1,200 in + 400 out tokens)$0.0003
1,000 requests / month$0.264/mo10,000 requests / month$2.64/mo100,000 requests / month$26.40/moEstimate your own workloadOpens the calculator with DeepSeek V4 Flash 0731 (batch) preloaded — adjust volume and token counts there.
Price history
List price per 1M tokens since we started tracking in Aug 2026. Each step marks a price change — hover a point for the exact date and rate.
What it excels at
Handles large-context batch workloads at very low cost. Performs well on coding and math tasks relative to price.
The business tradeoff
Reasoning depth trails flagship closed models. Output quality and latency depend on deployment environment.
Cheaper from DeepSeek
Head-to-head
Vendor list rates, as of Sep 8, 2026 · source: openrouter · per-request examples assume 1,200 input + 400 output tokens.