← All open-weight models

DeepSeek · General language

DeepSeek-V4-Pro

DeepSeek's frontier tier — worth the 3x premium over Flash only for hard agentic/production coding, deep reasoning and tool-heavy workflows where Flash's quality gap actually shows; note the tighter 500-concurrent limit and use reasoning_effort=max sparingly since thinking tokens bill at the output rate.

Open weights · Permissive (Apache, MIT, OpenMDW) · Data checked September 6, 2026 · Not a hands-on evaluation

Model ID deepseek-v4-pro

Parameters
1.7T total · 49B active1.7T total MoE (~49B active; V4 preview quoted 1.6T total / 49B active)
Architecture
Mixture of experts
Context window
1,000,000 tokensMaximum input
Max output
384,000 tokensPer response
Input price
$1.32 per 1M tokens
Output price
$3.96 per 1M tokens
Cached input
$0.044 per 1M tokens
Blended price
$1.98 per 1M tokens3:1 input to output
Knowledge cutoff
Not recorded
Released
2026-08-13
Status
gaFlagship in its lineup
Size band
120B and above

License and openness

DeepSeek-V4-Pro is released under the MIT. This is a permissive license. It generally allows use, modification and redistribution, including commercial use, subject to attribution and notice requirements. License terms can change between versions; confirm the text published with the weights you download.

Capabilities

Input modalitiestext
Output modalitiestext

What it is for

At a three-to-one input-to-output ratio, DeepSeek-V4-Pro costs $1.98 per million tokens blended. A workload of one million input and 330,000 output tokens per day would run about $78.80 per month at list price, before caching or batch discounts.

Cost at list price

Input$1.32 per 1M tokens
Output$3.96 per 1M tokens
Cached input$0.044 per 1M tokens
Blended, 3:1$1.98 per 1M tokens
Estimate a workload across all models ↗

Sources

Sources checked September 6, 2026. Specifications and prices change without notice; confirm against the provider before you commit.

Other open-weight models from DeepSeek

Browse all open-weight models ↗