DeepSeek
DeepSeek V4-Pro
DeepSeek · frontier-open. Frontier-grade open-weight coding.
The published record
Every number below was read from the provider's own documentation on the date shown. benchr does not restate a figure it has not seen published.
- API identifier
deepseek-v4-pro- Context window
- 1MtokensSource
- Maximum output
- 384KtokensSource
- Input
- $1.32per 1M tokensSource
- Output
- $3.96per 1M tokensSource
- Released
- April 24, 2026Source
- License
- MIT
Record verified August 28, 2026
What the record says
UPDATED 2026-08-28: DeepSeek raised API prices and introduced peak/off-peak billing effective 16:00 UTC on August 16, 2026. Cache-miss input went from a flat $0.435 to $1.32 peak / $0.66 off-peak per 1M tokens, output from $0.87 to $3.96 / $1.98, and cache-hit input from $0.003625 to $0.044 / $0.022 - the steepest line on the sheet at roughly twelve times the old cache-hit rate. Peak hours are 01:00-04:00 and 06:00-10:00 UTC, Monday to Friday. DeepSeek also shipped a V4-Pro GA update on August 13, 2026 with a new provider-reported benchmark table (Terminal Bench 2.1 87.9, HLE 42.7 without tools / 60.0 with tools, NL2Repo 61.5, CyberGym 83.3, DeepSWE 62.7, Toolathlon-Verified 74.1, DSBench-Hard 67.2); those are DeepSeek's own numbers, not benchr tests. The April model-card figures are retained as published. Open weights, MIT.
What has moved
Entries in the change ledger that affect this model, newest first.
Price changed
Cache-miss input rose from a flat $0.435 to $1.32 peak / $0.66 off-peak per 1M tokens, output from $0.87 to $3.96 / $1.98, and cache-hit input from $0.003625 to $0.044 / $0.022. DeepSeek replaced flat API pricing with peak/off-peak billing at 16:00 UTC on August 16, 2026. Peak hours are 01:00-04:00 and 06:00-10:00 UTC, Monday to Friday; all other hours bill at half the peak rate. The pricing figures here are the peak (list) rates. Source: api-docs.deepseek.com/quick_start/pricing (read 2026-08-28).
Released
DeepSeek V4-Pro launched April 24, 2026. MIT license. Source: api-docs.deepseek.com/quick_start/pricing
Which one to use
The rest of the family, with the two numbers that usually decide it. The current model is marked.
| Models | Input | Output | Context window |
|---|---|---|---|
| DeepSeek V4-Flash | $0.44 | $1.32 | 1M |
| DeepSeek V4-Pro This page | $1.32 | $3.96 | 1M |
benchr's read
Worth it for
- Frontier-grade open-weight coding
- Math-heavy work
- Self-hosted production, where the August 2026 API price rise does not apply
Look elsewhere if
- You need vision/multimodal
- You can't manage GPU hosting
Prices, limits, and identifiers above are provider facts. This section is benchr's judgement about them.
What benchr has written about it
Pieces that name this model, newest first. The ones written about this model come before the ones that mention it in passing.
What is not on this page
Stated rather than filled in.
- No first-token or tokens-per-second figure is recorded. benchr has not measured it and the provider does not publish one.