Skip to content

DeepSeek V4 Pro

DeepSeek · Released 24 Apr 2026 · deepseek-v4-pro

Compare this model
Open weightCodingAgentsReasoningLong context

$1.32

Input price / 1M tokens

#31 of 61, cheapest first

$3.96

Output price / 1M tokens

#24 of 61, cheapest first

1.05M

Context window

#11 of 71

104

Output tokens / second

#28 of 61

87.9%

Terminal-Bench 2.1

#6 of 17

42.7%

Humanity’s Last Exam (no tools)

#7 of 23

60%

Humanity’s Last Exam (with tools)

#8 of 16

62.7%

DeepSWE v1.1

#15 of 17

Summary

DeepSeek V4 Pro is an open-weight model from DeepSeek, released on 24 Apr 2026.

At $1.32 input and $3.96 output per million tokens, it is #31 of 61 on input price, cheapest first.

Its 1.05M-token context window ranks #11 of 71.

Measured output speed is 104 tokens per second, #28 of 61.

DeepSeek's flagship MoE model, 1.6T total and 49B active parameters, open weights under MIT (commercial use allowed, no extra conditions). The API now serves the GA build DeepSeek-V4-Pro-0813 (13 Aug 2026), which DeepSeek says greatly improves agent performance and adds low/high/max reasoning effort.

Benchmarks · 6 reported

BenchmarkScoreBarRankSource
Terminal-Bench 2.1V4-Pro-0813, DeepSeek Harness minimal mode, max effort87.9%#6 of 17DeepSeek-V4-Pro-0813 model cardSelf-reported
Humanity’s Last Exam (no tools)no tools, text-only subset42.7%#7 of 23DeepSeek-V4-Pro-0813 model cardSelf-reported
Humanity’s Last Exam (with tools)with tools60%#8 of 16DeepSeek-V4-Pro-0813 model cardSelf-reported
DeepSWE v1.1max effort62.7%#15 of 17DeepSeek-V4-Pro-0813 model cardSelf-reported
Toolathlon-Verified74.1%–DeepSeek-V4-Pro-0813 model cardSelf-reported
GPQA DiamondPass@1, max effort92.4%#9 of 27DeepSeek-V4.1-Flash model card comparison tableSelf-reported

Self-reported means the lab ran the test itself. A rank appears only where several labs report the same version of a benchmark.

Similar models

Models with the most benchmarks in common with DeepSeek V4 Pro, and the closest scores.

Compare 4 side by side

Terminal-Bench 2.1

#6 of 17

Terminal tasks, earlier set

Humanity’s Last Exam (no tools)

#7 of 23

Expert questions, no search or code

Humanity’s Last Exam (with tools)

#8 of 16

Expert questions, with search and code

DeepSWE v1.1

#15 of 17

Long coding tasks in real repos

Compare DeepSeek V4 Pro side by side

Pre-filled with the two closest models. Swap any of them.

Open comparison

Sources

Building on DeepSeek V4 Pro?

Get the architecture right before the bill arrives.

We size caching, routing and fallbacks for your workload on a free call, and tell you where a cheaper model is good enough.