The V4 generation: 1M-token context, 384K max output, and cache hits at ~3% of fresh input cost — with 50% off-peak discounts.
2 models tracked · cheapest input: DeepSeek V4 Flash at $0.44/M · official pricing page
deepseek-v4-flash
Input /M
$0.44
Output /M
$1.32
Context
1M
The value monster of 2026: $0.44/$1.32 with a 1M-token context and 384K max output, cache hits at $0.014/M.
deepseek-v4-pro
Input /M
$1.32
Output /M
$3.96
Context
1M
DeepSeek's frontier tier at $1.32/$3.96 peak — 1M context, 384K output, thinking and non-thinking modes.
Other providers
All DeepSeek prices verified Aug 16, 2026. Prices change frequently — each model page links to the authoritative source.