System Card
← leaderboard
Mistral
Mistral
model systemcard
Mistral Large 3
#27
of 27 · general
675B
params · MoE
Apache 2.0
license
2025-09-01
released
Standing
dot = this model · shaded = the field's middle 60% · track = full field
General
5.4
#27
/27
Math
30.6
#20
/27
Coding
19.0
#24
/27
Health
61.8
#9
/26
Legal
55.6
#7
/26
Instruct
2.3
#26
/27
Agents
2.2
#27
/27
Professional
2.5
#26
/27
Multimodal
5.2
#18
/19
Hallucination
10.5
#24
/27
Speed
38.3
#15
/22
Value
22.2
#24
/27
Open
19.8
#10
/11
Measured
≈ imputed · amber dot = lab-self-reported among the sources
GPQA Diamond
68.0%
#26
HLE
4.2%
#25
MMLU-Pro
79.8%
#24
AIME 2025
42.9%
#9
SWE-bench V
41.4%
#27
LiveCodeBench
44.9%
#24
SciCode
36.2%
#26
CritPt
0.0%
#25
Terminal-Bench
10.5%
#27
GDPval
7.0%
#22
Arena
1415
#24
AA Index
15.9
#23
Operations
$0.50
/
$1.5
$ per 1M tokens, in / out
$0.75
blended
39
tokens / sec
1415
arena elo · 52,509 votes
Series
Mistral's models on the board, by release
Mistral Large 3
2025-09-01 · 5
→
Mistral Medium
2026-04-29 · 9
Neighbors
nearest capability fingerprints · deltas describe the neighbor
Mistral
Mistral Medium
→ follows instructions better · pricier
MiniMax
MiniMax M2.5
→ weaker math · stronger coding
Tencent
HY3
→ stronger agent · cheaper
MiniMax
MiniMax M2.7
→ weaker math
Moonshot (月之暗面)
Kimi K2
→ stronger coding · follows instructions better