System Card
← leaderboard
OpenAI
model systemcard
GPT-5.5
#6
of 27 · general
incl. pre-release runs; Pro is a different product and excluded
1.1M
context
Proprietary
non-commercial
2026-04-23
released
2025-12-01
knowledge cutoff
Standing
dot = this model · shaded = the field's middle 60% · track = full field
General
67.2
#6
/27
Math
74.2
#5
/27
Coding
62.7
#8
/27
Health
40.2
#15
/26
Cyber
84.6
#3
/14
Analyst
90.0
#2
/11
Legal
30.9
#20
/26
Instruct
67.4
#5
/27
Agents
48.9
#13
/27
Professional
54.9
#10
/27
Multimodal
64.2
#4
/19
Hallucination
59.6
#2
/27
Value
55.7
#9
/27
Measured
≈ imputed · amber dot = lab-self-reported among the sources
GPQA Diamond
93.6%
#4
HLE
52.2%
#3
MMLU-Pro
88.1%
#10
LiveBench
80.4%
#3
SWE-bench V
78.4%
#14
LiveCodeBench
85.3%
#15
SciCode
56.1%
#6
CritPt
27.1%
#5
Terminal-Bench
78.0%
#8
APEX Agents
38.5%
#7
Arena
1482
#8
Vals Index
68.0
#8
Epoch Capabilities Index
159.2
#4
Operations
$0.43
$ per solved task
1482
arena elo · 47,572 votes
Series
OpenAI's models on the board, by release
GPT-5.5
2026-04-23 · 67
→
GPT-5.6 Sol
2026-07-09 · 85
→
GPT-5.6 Terra
2026-07-09 · 59
→
GPT-5.6 Luna
2026-07-09 · 50
Neighbors
nearest capability fingerprints · deltas describe the neighbor
GPT-5.6 Sol
→ stronger coding
Zhipu (智谱)
GLM-5.2
→ looser instructions
DeepMind
Gemini 3.5 Flash
→ weaker coding
Moonshot (月之暗面)
Kimi K3
→ weaker math
Alibaba
Qwen3.8 Max
→ weaker math · weaker multimodal