System Card
← leaderboard
OpenAI
model systemcard
GPT-5.6 Luna
#13
of 27 · general
effort ladder collapsed
1.1M
context
Proprietary
non-commercial
2026-07-09
released
2026-02-16
knowledge cutoff
Standing
dot = this model · shaded = the field's middle 60% · track = full field
General
49.7
#13
/27
Math
41.6
#15
/27
Coding
39.0
#15
/27
Health
27.3
#20
/26
Engineering
31.2
#12
/17
Cyber
92.3
#2
/14
Legal
66.8
#5
/26
Instruct
3.1
#25
/27
Agents
64.2
#8
/27
Professional
42.2
#15
/27
Multimodal
38.6
#12
/19
Hallucination
19.3
#21
/27
Speed
97.4
#1
/22
Value
57.8
#7
/27
Measured
≈ imputed · amber dot = lab-self-reported among the sources
GPQA Diamond
91.6%
#13
HLE
39.5%
#17
MMLU-Pro
86.0%
#20
LiveBench
73.6%
#13
SWE-bench V
93.0%
#5
SciCode
52.5%
#15
CritPt
20.6%
#8
Terminal-Bench
80.0%
#7
APEX Agents
35.8%
#8
GDPval
54.0%
#8
AA Index
52.3
#11
Vals Index
69.9
#5
Epoch Capabilities Index
156.1
#7
Operations
$0.20
/
$1.2
$ per 1M tokens, in / out
$0.45
blended
195
tokens / sec
$0.17
$ per solved task
~1477
arena elo · predicted, never voted on
Series
OpenAI's models on the board, by release
GPT-5.5
2026-04-23 · 67
→
GPT-5.6 Sol
2026-07-09 · 85
→
GPT-5.6 Terra
2026-07-09 · 59
→
GPT-5.6 Luna
2026-07-09 · 50
Neighbors
nearest capability fingerprints · deltas describe the neighbor
GPT-5.6 Terra
→ stronger math · pricier
xAI
Grok 4.5
→ follows instructions better · pricier
Tencent
HY3
→ weaker agent · cheaper
Anthropic
Claude Sonnet 5
→ follows instructions better · pricier
Moonshot (月之暗面)
Kimi K2
→ near-twin