DeepSeek · General language
DeepSeek-V4-Pro
DeepSeek's frontier tier — worth the 3x premium over Flash only for hard agentic/production coding, deep reasoning and tool-heavy workflows where Flash's quality gap actually shows; note the tighter 500-concurrent limit and use reasoning_effort=max sparingly since thinking tokens bill at the output rate.
Open weights · Permissive (Apache, MIT, OpenMDW) · Data checked September 6, 2026 · Not a hands-on evaluation
Model ID deepseek-v4-pro
- Parameters
- 1.7T total · 49B active1.7T total MoE (~49B active; V4 preview quoted 1.6T total / 49B active)
- Architecture
- Mixture of experts
- Context window
- 1,000,000 tokensMaximum input
- Max output
- 384,000 tokensPer response
- Input price
- $1.32 per 1M tokens
- Output price
- $3.96 per 1M tokens
- Cached input
- $0.044 per 1M tokens
- Blended price
- $1.98 per 1M tokens3:1 input to output
- Knowledge cutoff
- Not recorded
- Released
- 2026-08-13
- Status
- gaFlagship in its lineup
- Size band
- 120B and above
License and openness
DeepSeek-V4-Pro is released under the MIT. This is a permissive license. It generally allows use, modification and redistribution, including commercial use, subject to attribution and notice requirements. License terms can change between versions; confirm the text published with the weights you download.
Capabilities
| Input modalities | text |
|---|---|
| Output modalities | text |
What it is for
At a three-to-one input-to-output ratio, DeepSeek-V4-Pro costs $1.98 per million tokens blended. A workload of one million input and 330,000 output tokens per day would run about $78.80 per month at list price, before caching or batch discounts.
Cost at list price
| Input | $1.32 per 1M tokens |
|---|---|
| Output | $3.96 per 1M tokens |
| Cached input | $0.044 per 1M tokens |
| Blended, 3:1 | $1.98 per 1M tokens |
Sources
- Source: DeepSeek pricing ↗
- API docs ↗
- Full reference record for DeepSeek-V4-Pro ↗
- DeepSeek in the reference ↗
Sources checked September 6, 2026. Specifications and prices change without notice; confirm against the provider before you commit.
Other open-weight models from DeepSeek
- DeepSeek-V4-Flash · 304B total · 13B active
- DeepSeek-V4-Flash-Vision-Exp · 305B