System Card
← leaderboard
Anthropic
Anthropic
model systemcard
Claude Opus 5
#2
of 27 · general
max/high effort collapsed
1M
context
Proprietary
non-commercial
2026-07-24
released
paper ↗
Standing
dot = this model · shaded = the field's middle 60% · track = full field
General
89.2
#2
/27
Math
86.8
#2
/27
Coding
78.7
#3
/27
Health
86.2
#2
/26
Engineering
100.0
#1
/17
Cyber
7.7
#13
/14
Analyst
100.0
#1
/11
Legal
67.4
#4
/26
Instruct
51.5
#11
/27
Agents
97.4
#1
/27
Professional
100.0
#1
/27
Multimodal
91.7
#1
/19
Hallucination
55.7
#3
/27
Speed
30.5
#17
/22
Value
68.7
#3
/27
Measured
≈ imputed · amber dot = lab-self-reported among the sources
GPQA Diamond
93.7%
#3
HLE
59.8%
#2
MMLU-Pro
91.6%
#1
LiveBench
80.1%
#4
SWE-bench V
97.0%
#1
LiveCodeBench
89.0%
#4
SciCode
55.7%
#7
CritPt
29.1%
#3
Terminal-Bench
86.9%
#2
APEX Agents
43.5%
#2
GDPval
67.4%
#1
Arena
1495
#2
AA Index
63.0
#1
Vals Index
74.8
#2
Epoch Capabilities Index
161.0
#3
Operations
$5
/
$25
$ per 1M tokens, in / out
$10
blended
53
tokens / sec
$0.70
$ per solved task
1495
arena elo · 2,386 votes
Series
Anthropic's models on the board, by release
Claude Fable 5
2026-06-09 · 96
→
Claude Sonnet 5
2026-06-30 · 60
→
Claude Opus 5
2026-07-24 · 89
Neighbors
nearest capability fingerprints · deltas describe the neighbor
Anthropic
Claude Fable 5
→ follows instructions better · pricier
Muse Spark
→ weaker math · cheaper
GPT-5.6 Sol
→ weaker agent
Moonshot (月之暗面)
Kimi K3
→ cheaper
Zhipu (智谱)
GLM-5.2
→ cheaper