deepseek-v4-pro · DeepSeek
Next-gen MoE frontier model with 1M context, 384k max output, and dynamic off-peak discounts ($0.66/$1.98 off-peak, $1.32/$3.96 peak).
Last checked Aug 28, 2026. Calculations use the listed base rates; provider-specific tiers, cache writes, batch pricing, and long-context rules may differ. Note: Off-peak hours: $0.66/1M input, $1.98/1M output. Cache hit rate is $0.022 off-peak / $0.044 peak. Confirm the current rate at the official provider source.
Quick answer
At the published base rate, DeepSeek V4 Pro costs $1.32 per 1M input tokens and $3.96 per 1M output tokens. It supports a 1M context window and is currently marked ga.
Method & trust
This rate card uses the provider's listed base input, output, and cached-input prices. Workload tables below apply those rates to explicit token counts, cache assumptions, and request volumes; verify provider-specific tiers before committing budget.
Input
$1.32
per 1M tokens
Output
$3.96
per 1M tokens
Cached input
$0.044
per 1M tokens
Context window
1M
max output 384K
DeepSeek V4 Pro cost calculator
What DeepSeek V4 Pro costs per task
| Use case | Input tokens | Output tokens | Cost / request | Monthly @ 1K req/day |
|---|---|---|---|---|
| Customer support chatbot | 3,500 | 350 | $0.00288 | $86.394 |
| RAG / search-augmented answers | 8,000 | 500 | $0.007436 | $223.08 |
| AI coding assistant | 12,000 | 2,000 | $0.0146 | $437.184 |
| Document summarization | 25,000 | 600 | $0.0322 | $965.58 |
| Agentic workflow | 40,000 | 1,500 | $0.0256 | $766.92 |
| Content generation | 800 | 1,200 | $0.005502 | $165.053 |
| Data extraction & tagging | 2,000 | 250 | $0.00312 | $93.588 |
| Translation | 5,000 | 5,500 | $0.0274 | $822.69 |
Assumes each use case's typical cacheable share of input. See full cost scenarios.
Benchmark scores
Third-party composite indices (0-100) via OpenRouter's model catalog, verified Aug 28, 2026. Coding per $$3.96 of output: 17.4.
Compare with alternatives
| Model | Input /M | Output /M | Context | Chat request* |
|---|---|---|---|---|
| DeepSeek V4 Pro | $1.32 | $3.96 | 1M | $0.005896 |
| GPT-5.6 Solcompare | $5.00 | $30.00 | 1.1M | $0.035 |
| GPT-5.6 Terracompare | $2.00 | $12.00 | 1.1M | $0.014 |
| GPT-5.6 Lunacompare | $0.20 | $1.20 | 1.1M | $0.0014 |
| GPT-5.6 Cybercompare | $12.50 | $75.00 | 1.1M | $0.0875 |
| GPT-5.5 Standardcompare | $5.00 | $30.00 | 512K | $0.035 |
| GPT-5.5 Procompare | $30.00 | $180.00 | 512K | $0.21 |
| GPT-5.4 Workhorsecompare | $2.50 | $15.00 | 256K | $0.0175 |
| GPT-5.4 minicompare | $0.75 | $4.50 | 256K | $0.00525 |
*4,000 in + 800 out tokens, 50% cached input where available.
Related calculations
FAQ
DeepSeek V4 Pro costs $1.32 per 1M input tokens and $3.96 per 1M output tokens, with cached input at $0.044 per 1M tokens. Verified Aug 28, 2026 against https://api-docs.deepseek.com/quick_start/pricing.
DeepSeek V4 Pro supports a 1,000,000-token context window with up to 384,000 output tokens per request, tokenized with deepseek_bpe.
At $1.32/M input and $3.96/M output, DeepSeek V4 Pro sits below GPT-5.6 Sol ($5.00/M in, $30.00/M out) — see the comparison table for full-workload differences.