DeepSeek-V4-Flash
The default DeepSeek pick: near-Pro reasoning at a third of the price with the same 1M context, ideal for high-volume agentic coding and long-document work — schedule batchy jobs outside 01:00-04:00/06:00-10:00 UTC weekdays and you pay $0.22/$0.66 instead of $0.44/$1.32.
Data checked 2026-09-06
DeepSeek
ga
open weights
deepseek-v4-flash
- Context window
- 1M tokens in
- Max output
- 384K tokens out
- Input
- $0.44 per 1M tokens
- Output
- $1.32 per 1M tokens
- Cached input
- $0.01 per 1M tokens
- Blended
- $0.66 3:1 in:out
- Released
- 2026-07-31
- Parameters
- 304B total MoE (~13B active; V4 preview quoted 284B total / 13B active, architecture unchanged in the 0731 re-post-train)
- Licence
- MIT
A What it is for
At a three-to-one input-to-output ratio, DeepSeek-V4-Flash costs $0.66 per million tokens blended. A workload of one million input and 330,000 output tokens per day would run about $26.27 per month at list price, before caching or batch discounts.
B Capabilities
| Input modalities | text |
|---|---|
| Output modalities | text |
C Others from DeepSeek
| Model | Context | Max out | In $/M | Out $/M | Blended | Status |
|---|---|---|---|---|---|---|
DeepSeek-V4-Pro
deepseek-v4-pro
|
1M | 384K | $1.32 | $3.96 | $1.98 | ga open |
DeepSeek-V4-Flash-Vision-Exp
deepseek-v4-flash-vision-exp
|
1M | 384K | $0.44 | $1.32 | $0.66 | preview open |
D Verify before you commit
Model pricing changes without notice and this page is a snapshot. Confirm against the vendor's own page before you build a budget on it.
Source: api-docs.deepseek.com · checked 2026-09-06 · DeepSeek pricing · API docs