DeepSeek API pricing: 2 models, per token and per task
The list price of every DeepSeek model on our list, read from DeepSeek’s own pricing page, next to what a passing answer cost on the AIGROW small-business benchmark and what three everyday jobs cost on each.
US dollars per million tokens, at the standard rate. No reseller prices.
- DeepSeek V4 Pro
- $1.32 / $3.96
- DeepSeek V4.1 Flash
- $0.30 / $1.20
Input / output list price per million tokens.
DeepSeek prices the 2 models on our list from DeepSeek V4.1 Flash at $0.30 in and $1.20 out per million tokens to DeepSeek V4 Pro at $1.32 in and $3.96 out. On the AIGROW small-business benchmark DeepSeek V4.1 Flash shared the best record, 18 of 20 tasks passed, at the lowest cost: $1.05 per thousand passes.
Per token and per pass
The list price per million tokens, then how many of the benchmark’s 20 office tasks each model passed and what 1,000 passing answers cost, failed runs included.
| Model | Input | Output | Cached input | Context | Tasks passed | 1,000 passes |
|---|---|---|---|---|---|---|
| DeepSeek V4 ProFlagships | $1.32 | $3.96 | $0.044 | 1M | 18 of 20 | $5.14 |
| DeepSeek V4.1 FlashSmall and fast | $0.30 | $1.20 | $0.006 | 1M | 18 of 20 | $1.05 |
What three everyday jobs cost
1,000 runs of each job at list price. The token counts are ours; the prices are DeepSeek’s.
| Model | 1,000 × support reply | 1,000 × document summary | 1,000 × agent task |
|---|---|---|---|
| DeepSeek V4 Pro | $5.54 | $36.96 | $28.16 |
| DeepSeek V4.1 Flash | $1.38 | $8.70 | $7.02 |
- Support reply
- A 3,000-token prompt (the ticket, the thread so far and a policy excerpt) and a 400-token answer.
- Document summary
- A 40-page document of about 25,000 tokens in, a 1,000-token summary out.
- Agent task
- Ten steps that read 80,000 tokens between them, 70,000 of them context repeated from step to step and served from cache, and write 3,000.
The small print
What DeepSeek’s pricing pages add to the headline rates.
- DeepSeek V4 Pro
- Peak-hour rate, recorded as the list price. Off-peak (all hours outside 01:00-04:00 and 06:00-10:00 UTC on weekdays) is half: $0.66 / $1.98, cache hit $0.022. Peak and off-peak pricing started 2026-08-16. API name deepseek-v4-pro serves DeepSeek-V4-Pro-0813 (V4 first shipped 2026-04-24). Max output 384K.
- DeepSeek V4.1 Flash
- Peak-hour rate, recorded as the list price. Off-peak is half: $0.15 / $0.60, cache hit $0.003. API name deepseek-flash. Max output 384K.
Head to head
The comparisons with a DeepSeek model on one side, each with both prices and the three jobs costed on each.
- DeepSeek V4 Pro vs Claude Opus 5
- DeepSeek V4 Pro vs GPT-5.6 Sol
- Qwen3.8 Max vs DeepSeek V4 Pro
- GLM-5.3 vs DeepSeek V4 Pro
- MiniMax M3 vs DeepSeek V4 Pro
- gpt-oss-120b vs DeepSeek V4.1 Flash
- DeepSeek V4.1 Flash vs GPT-5.6 Luna
- DeepSeek V4.1 Flash vs Gemini 3.5 Flash-Lite
- Qwen3.8 Flash vs DeepSeek V4.1 Flash
- Grok 4.3 vs DeepSeek V4.1 Flash
Where the prices come from
Read on 22 September 2026. Each price is the standard rate for the shortest context band, with no batch or priority discount; long-context surcharges, batch rates and promotions are in each model’s small print. No price was taken from a reseller.
Pay a quarter of list price
AIGROW API credit is metered at the makers’ list prices, like the ones above, and $25 buys $100 of usage. The credit page lists the models it covers.
Prices checked 22 September 2026