deepseek-chat · DeepSeek
671B parameter MoE architecture (37B active) delivering frontier chat performance at $0.14/$0.28 per 1M tokens.
Last checked Aug 19, 2026. Calculations use the listed base rates; provider-specific tiers, cache writes, batch pricing, and long-context rules may differ. Note: Cache hits bill at 90% discount ($0.014/1M). Off-peak hours offer 50% discount. Confirm the current rate at the official provider source.
Input
$0.14
per 1M tokens
Output
$0.28
per 1M tokens
Cached input
$0.014
per 1M tokens
Context window
64K
max output 8K
DeepSeek V3 (Chat) cost calculator
What DeepSeek V3 (Chat) costs per task
| Use case | Input tokens | Output tokens | Cost / request | Monthly @ 1K req/day |
|---|---|---|---|---|
| Customer support chatbot | 3,500 | 350 | $0.000279 | $8.379 |
| RAG / search-augmented answers | 8,000 | 500 | $0.000756 | $22.68 |
| AI coding assistant | 12,000 | 2,000 | $0.001333 | $39.984 |
| Document summarization | 25,000 | 600 | $0.003353 | $100.59 |
| Agentic workflow | 40,000 | 1,500 | $0.002744 | $82.32 |
| Content generation | 800 | 1,200 | $0.000418 | $12.533 |
| Data extraction & tagging | 2,000 | 250 | $0.0003 | $8.988 |
| Translation | 5,000 | 5,500 | $0.002146 | $64.365 |
Assumes each use case's typical cacheable share of input. See full cost scenarios.
Compare with alternatives
| Model | Input /M | Output /M | Context | Chat request* |
|---|---|---|---|---|
| DeepSeek V3 (Chat) | $0.14 | $0.28 | 64K | $0.000532 |
| GPT-5.6 Sol | $5.00 | $30.00 | 1.1M | $0.035 |
| GPT-5.6 Terra | $2.00 | $12.00 | 1.1M | $0.014 |
| GPT-5.6 Luna | $0.20 | $1.20 | 1.1M | $0.0014 |
| GPT-5.6 Cyber | $12.50 | $75.00 | 1.1M | $0.0875 |
| GPT-5.5 Standard | $5.00 | $30.00 | 512K | $0.035 |
| GPT-5.5 Pro | $30.00 | $180.00 | 512K | $0.21 |
| GPT-5.4 Workhorse | $2.50 | $15.00 | 256K | $0.0175 |
| GPT-5.4 mini | $0.75 | $4.50 | 256K | $0.00525 |
*4,000 in + 800 out tokens, 50% cached input where available.
Related calculations
FAQ
DeepSeek V3 (Chat) costs $0.14 per 1M input tokens and $0.28 per 1M output tokens, with cached input at $0.014 per 1M tokens. Verified Aug 19, 2026 against https://api-docs.deepseek.com/quick_start/pricing.
DeepSeek V3 (Chat) supports a 64,000-token context window with up to 8,000 output tokens per request, tokenized with deepseek_bpe.
At $0.14/M input and $0.28/M output, DeepSeek V3 (Chat) sits below GPT-5.6 Sol ($5.00/M in, $30.00/M out) — see the comparison table for full-workload differences.