System Card
← leaderboard
DeepSeek-AI
model systemcard
DeepSeek V4 Flash
#20
of 27 · general
284B
params · MoE
1M
context
32T
training tokens
MIT
license
2026-04-23
released
weights ↗
repo ↗
Standing
dot = this model · shaded = the field's middle 60% · track = full field
General
20.4
#20
/27
Math
6.8
#27
/27
Coding
16.0
#25
/27
Health
21.0
#21
/26
Engineering
25.0
#13
/17
Cyber
30.8
#10
/14
Legal
25.2
#22
/26
Instruct
20.3
#22
/27
Agents
33.6
#15
/27
Professional
37.9
#18
/27
Hallucination
12.6
#23
/27
Speed
66.7
#5
/22
Value
40.3
#19
/27
Open
44.2
#3
/11
Measured
≈ imputed · amber dot = lab-self-reported among the sources
GPQA Diamond
90.8%
#14
HLE
41.8%
#14
MMLU-Pro
86.2%
#19
LiveBench
65.5%
#20
SWE-bench V
79.0%
#11
LiveCodeBench
91.6%
#1
SciCode
49.9%
#19
CritPt
16.6%
#11
Terminal-Bench
67.8%
#16
GDPval
52.9%
#10
Arena
1438
#20
AA Index
51.8
#12
Epoch Capabilities Index
152.5
#15
Operations
$0.14
/
$0.28
$ per 1M tokens, in / out
$0.18
blended
124
tokens / sec
$0.02
$ per solved task
1438
arena elo · 46,445 votes
Series
DeepSeek's models on the board, by release
DeepSeek V4 Pro
2026-04-23 · 23
→
DeepSeek V4 Flash
2026-04-23 · 20
Neighbors
nearest capability fingerprints · deltas describe the neighbor
Tencent
HY3
→ more factual · weaker professional
TM
Inkling
→ stronger math · pricier
MiniMax
MiniMax M3
→ pricier
DeepSeek V4 Pro
→ stronger math · pricier
Nemotron 3 Ultra
→ pricier