LLMPrice
Updated daily
Direct API pricing comparison

DeepSeek V4 Pro vs DeepSeek V4 Flash

Compare current direct API token prices and estimated monthly workload costs for DeepSeek V4 Pro and DeepSeek V4 Flash.

2direct models
Pricing only.DeepSeek V4 Flash is cheaper in 4 of 4 comparable standard workload presets below. This is a pricing comparison only; it does not rank model quality.
Published direct rates

Token pricing side by side

USD per 1M text tokens
ModelInputCachedOutputContextMax outputSource
DeepSeek V4 ProDeepSeek · Direct API
$1.32$0.044$3.961M384KDirect
DeepSeek V4 FlashDeepSeek · Direct API
$0.44$0.014$1.321M384KDirect
Monthly estimates

Same workload, different API bill

These four presets use 30 active days, standard direct rates, no retries, no reasoning-token add-on, no cache writes, and the cached-input percentage shown in each row.

Standard workloadDeepSeek V4 ProDeepSeek V4 FlashLower monthly estimate
Chat10,000/day · 2,000 in · 500 out · 25% cached$1,194.6$398.1DeepSeek V4 Flash by $796.5 / mo
Coding2,000/day · 12,000 in · 1,800 out · 55% cached$872.78$290.66DeepSeek V4 Flash by $582.12 / mo
Agent1,200/day · 32,000 in · 2,500 out · 65% cached$921.57$306.69DeepSeek V4 Flash by $614.88 / mo
Documents400/day · 180,000 in · 2,000 out · 15% cached$2,532.82$844.06DeepSeek V4 Flash by $1,688.76 / mo
Custom workload

Use your own traffic

The main calculator includes batch mode, long-context rules, retries, reasoning tokens, cache writes, and other model-specific pricing logic that a static comparison cannot fully capture.

Related comparisons

Compare nearby alternatives