DeepSeek

DeepSeek V4-Pro

DeepSeek · frontier-open. Frontier-grade open-weight coding.

The published record

Every number below was read from the provider's own documentation on the date shown. benchr does not restate a figure it has not seen published.

API identifier
deepseek-v4-pro
Context window
1MtokensSource
Maximum output
384KtokensSource
Input
$1.32per 1M tokensSource
Output
$3.96per 1M tokensSource
Released
April 24, 2026Source
License
MIT

Record verified August 28, 2026

What the record says

UPDATED 2026-08-28: DeepSeek raised API prices and introduced peak/off-peak billing effective 16:00 UTC on August 16, 2026. Cache-miss input went from a flat $0.435 to $1.32 peak / $0.66 off-peak per 1M tokens, output from $0.87 to $3.96 / $1.98, and cache-hit input from $0.003625 to $0.044 / $0.022 - the steepest line on the sheet at roughly twelve times the old cache-hit rate. Peak hours are 01:00-04:00 and 06:00-10:00 UTC, Monday to Friday. DeepSeek also shipped a V4-Pro GA update on August 13, 2026 with a new provider-reported benchmark table (Terminal Bench 2.1 87.9, HLE 42.7 without tools / 60.0 with tools, NL2Repo 61.5, CyberGym 83.3, DeepSWE 62.7, Toolathlon-Verified 74.1, DSBench-Hard 67.2); those are DeepSeek's own numbers, not benchr tests. The April model-card figures are retained as published. Open weights, MIT.

What has moved

Entries in the change ledger that affect this model, newest first.

  1. Price changed

    Cache-miss input rose from a flat $0.435 to $1.32 peak / $0.66 off-peak per 1M tokens, output from $0.87 to $3.96 / $1.98, and cache-hit input from $0.003625 to $0.044 / $0.022. DeepSeek replaced flat API pricing with peak/off-peak billing at 16:00 UTC on August 16, 2026. Peak hours are 01:00-04:00 and 06:00-10:00 UTC, Monday to Friday; all other hours bill at half the peak rate. The pricing figures here are the peak (list) rates. Source: api-docs.deepseek.com/quick_start/pricing (read 2026-08-28).

    DeepSeek

  2. Released

    DeepSeek V4-Pro launched April 24, 2026. MIT license. Source: api-docs.deepseek.com/quick_start/pricing

    DeepSeek

The whole change ledger →

Which one to use

The rest of the family, with the two numbers that usually decide it. The current model is marked.

ModelsInput OutputContext window
DeepSeek V4-Flash$0.44$1.321M
DeepSeek V4-Pro This page$1.32$3.961M

benchr's read

Worth it for

  • Frontier-grade open-weight coding
  • Math-heavy work
  • Self-hosted production, where the August 2026 API price rise does not apply

Look elsewhere if

  • You need vision/multimodal
  • You can't manage GPU hosting

Prices, limits, and identifiers above are provider facts. This section is benchr's judgement about them.

What benchr has written about it

Pieces that name this model, newest first. The ones written about this model come before the ones that mention it in passing.

What is not on this page

Stated rather than filled in.

  • No first-token or tokens-per-second figure is recorded. benchr has not measured it and the provider does not publish one.