System Card
← leaderboard
OpenAI
model systemcard
GPT-5.6 Sol
#3
of 27 · general
xhigh/max effort collapsed
1.1M
context
Proprietary
non-commercial
2026-07-09
released
2026-02-16
knowledge cutoff
Standing
dot = this model · shaded = the field's middle 60% · track = full field
General
85.2
#3
/27
Math
81.1
#3
/27
Coding
90.0
#1
/27
Health
54.9
#11
/26
Engineering
87.5
#3
/17
Cyber
100.0
#1
/14
Analyst
70.0
#4
/11
Legal
55.4
#8
/26
Instruct
70.4
#4
/27
Agents
64.9
#7
/27
Professional
79.3
#5
/27
Multimodal
68.0
#3
/19
Hallucination
49.4
#7
/27
Speed
63.8
#6
/22
Value
66.4
#4
/27
Measured
≈ imputed · amber dot = lab-self-reported among the sources
GPQA Diamond
94.4%
#1
HLE
49.5%
#5
MMLU-Pro
89.1%
#8
LiveBench
81.0%
#1
SWE-bench V
96.2%
#2
LiveCodeBench
82.6%
#16
SciCode
56.9%
#5
CritPt
32.3%
#1
Terminal-Bench
87.6%
#1
APEX Agents
39.9%
#4
GDPval
61.4%
#4
Arena
1485
#6
AA Index
60.9
#3
Vals Index
73.1
#4
Epoch Capabilities Index
161.7
#1
Operations
$5
/
$30
$ per 1M tokens, in / out
$11
blended
64
tokens / sec
$0.44
$ per solved task
1485
arena elo · 8,359 votes
Series
OpenAI's models on the board, by release
GPT-5.5
2026-04-23 · 67
→
GPT-5.6 Sol
2026-07-09 · 85
→
GPT-5.6 Terra
2026-07-09 · 59
→
GPT-5.6 Luna
2026-07-09 · 50
Neighbors
nearest capability fingerprints · deltas describe the neighbor
Moonshot (月之暗面)
Kimi K3
→ weaker math · cheaper
GPT-5.5
→ weaker coding
Anthropic
Claude Fable 5
→ more factual · pricier
Muse Spark
→ weaker math · cheaper
Zhipu (智谱)
GLM-5.2
→ cheaper