System Card
← leaderboard
DeepSeek-AI
model systemcard
DeepSeek V4 Pro
#19
of 27 · general
thinking mode collapsed
1.6T
params · MoE
1M
context
32T
training tokens
MIT
license
2026-04-23
released
weights ↗
repo ↗
Standing
dot = this model · shaded = the field's middle 60% · track = full field
General
23.0
#19
/27
Math
31.9
#19
/27
Coding
21.7
#22
/27
Health
6.2
#24
/26
Cyber
38.5
#9
/14
Legal
12.4
#24
/26
Instruct
38.9
#14
/27
Agents
31.6
#16
/27
Professional
38.3
#17
/27
Hallucination
25.1
#15
/27
Speed
59.5
#7
/22
Value
39.3
#20
/27
Open
42.3
#4
/11
Measured
≈ imputed · amber dot = lab-self-reported among the sources
GPQA Diamond
90.3%
#18
HLE
42.9%
#11
MMLU-Pro
87.4%
#13
LiveBench
72.6%
#16
SWE-bench V
77.6%
#17
LiveCodeBench
90.5%
#2
SciCode
50.0%
#18
CritPt
12.9%
#15
Terminal-Bench
64.8%
#18
APEX Agents
24.3%
#15
GDPval
40.3%
#15
Arena
1457
#17
AA Index
45.3
#17
Vals Index
55.6
#16
Epoch Capabilities Index
149.1
#19
Operations
$0.43
/
$0.87
$ per 1M tokens, in / out
$0.54
blended
71
tokens / sec
$0.05
$ per solved task
1457
arena elo · 46,848 votes
Series
DeepSeek's models on the board, by release
DeepSeek V4 Pro
2026-04-23 · 23
→
DeepSeek V4 Flash
2026-04-23 · 20
Neighbors
nearest capability fingerprints · deltas describe the neighbor
TM
Inkling
→ weaker professional · pricier
Tencent
HY3
→ weaker math · cheaper
MiniMax
MiniMax M3
→ near-twin
DeepSeek V4 Flash
→ weaker math · cheaper
DeepMind
Gemini 3.5 Flash
→ follows instructions better · pricier