System Card
← leaderboard
OpenAI
model systemcard
GPT-5.6 Terra
#8
of 27 · general
effort ladder collapsed
1.1M
context
Proprietary
non-commercial
2026-07-09
released
2026-02-16
knowledge cutoff
Standing
dot = this model · shaded = the field's middle 60% · track = full field
General
59.2
#8
/27
Math
67.7
#7
/27
Coding
61.2
#9
/27
Health
31.1
#18
/26
Engineering
50.0
#9
/17
Legal
49.5
#13
/26
Instruct
28.9
#20
/27
Agents
59.0
#10
/27
Professional
39.2
#16
/27
Multimodal
40.3
#11
/19
Hallucination
18.2
#22
/27
Speed
84.3
#3
/22
Value
50.0
#11
/27
Measured
≈ imputed · amber dot = lab-self-reported among the sources
GPQA Diamond
92.7%
#8
HLE
42.9%
#10
MMLU-Pro
86.7%
#16
LiveBench
77.9%
#7
SWE-bench V
75.2%
#23
LiveCodeBench
85.9%
#12
SciCode
53.9%
#9
CritPt
30.0%
#2
Terminal-Bench
80.7%
#5
APEX Agents
38.9%
#6
GDPval
53.9%
#9
AA Index
56.6
#7
Vals Index
65.1
#10
Epoch Capabilities Index
159.0
#5
Operations
$2
/
$12
$ per 1M tokens, in / out
$4.5
blended
133
tokens / sec
$0.35
$ per solved task
~1484
arena elo · predicted, never voted on
Series
OpenAI's models on the board, by release
GPT-5.5
2026-04-23 · 67
→
GPT-5.6 Sol
2026-07-09 · 85
→
GPT-5.6 Terra
2026-07-09 · 59
→
GPT-5.6 Luna
2026-07-09 · 50
Neighbors
nearest capability fingerprints · deltas describe the neighbor
xAI
Grok 4.5
→ follows instructions better · cheaper
Anthropic
Claude Sonnet 5
→ stronger agent
GPT-5.6 Luna
→ weaker math · cheaper
Zhipu (智谱)
GLM-5.2
→ more factual · cheaper
Alibaba
Qwen3.8 Max
→ weaker math · cheaper