Seminal AI
§1

DeepSeek-V4-Flash

The default DeepSeek pick: near-Pro reasoning at a third of the price with the same 1M context, ideal for high-volume agentic coding and long-document work — schedule batchy jobs outside 01:00-04:00/06:00-10:00 UTC weekdays and you pay $0.22/$0.66 instead of $0.44/$1.32.

Data checked 2026-09-06

DeepSeek ga open weights deepseek-v4-flash

Context window
1M tokens in
Max output
384K tokens out
Input
$0.44 per 1M tokens
Output
$1.32 per 1M tokens
Cached input
$0.01 per 1M tokens
Blended
$0.66 3:1 in:out
Released
2026-07-31
Parameters
304B total MoE (~13B active; V4 preview quoted 284B total / 13B active, architecture unchanged in the 0731 re-post-train)
Licence
MIT

A What it is for

At a three-to-one input-to-output ratio, DeepSeek-V4-Flash costs $0.66 per million tokens blended. A workload of one million input and 330,000 output tokens per day would run about $26.27 per month at list price, before caching or batch discounts.

B Capabilities

tools reasoning prompt-caching streaming structured-output
Input modalitiestext
Output modalitiestext

C Others from DeepSeek

Model Context Max out In $/M Out $/M Blended Status
DeepSeek-V4-Pro deepseek-v4-pro 1M 384K $1.32 $3.96 $1.98 ga open
DeepSeek-V4-Flash-Vision-Exp deepseek-v4-flash-vision-exp 1M 384K $0.44 $1.32 $0.66 preview open

D Verify before you commit

Model pricing changes without notice and this page is a snapshot. Confirm against the vendor's own page before you build a budget on it.

Source: api-docs.deepseek.com · checked 2026-09-06 · DeepSeek pricing · API docs