← All models
Flash — fastest and cheapest, lighter reasoningOpen Weights
DeepSeek V4.1 Flash Pricing
DeepSeek · deepseek-deepseek-v4-1-flash
Input / 1M tokens
$0.150
Output / 1M tokens
$0.600
Context window
1,048,576 tokens
≈ 1,398 pages of text
Max output
384k tokens
What it costs in practice
Typical request (1,200 in + 400 out tokens)$0.0004
1,000 requests / month$0.42/mo10,000 requests / month$4.20/mo100,000 requests / month$42.00/moEstimate your own workloadOpens the calculator with DeepSeek V4.1 Flash preloaded — adjust volume and token counts there.
Price history
List price per 1M tokens since we started tracking in Sep 2026. Each step marks a price change — hover a point for the exact date and rate.
What it excels at
Strong performance on code generation and mathematical tasks with 1M-token context support at low per-token cost.
The business tradeoff
Output quality varies with self-hosted inference; generally trails closed flagships on open-ended reasoning and instruction adherence.
Cheaper from DeepSeek
Head-to-head
Vendor list rates, as of Sep 10, 2026 · source: openrouter · per-request examples assume 1,200 input + 400 output tokens.