← comparisons · head to head

GPT-5.6 vs Claude Sonnet 5

GPT-5.6 scores higher than Claude Sonnet 5 on all 5 public leaderboards that rate both. On the ArtificialWatch Index, which takes the median of each model's percentile across the boards that score it, GPT-5.6 is at 70.2 and Claude Sonnet 5 at 36.7 (out of 100).

Scores as published by each board; table composed September 25, 2026.

Get the next launch alert, free →

Leaderboard by leaderboard

BoardGPT-5.6Claude Sonnet 5Higher
LMArena
Elo (community votes)
1,617.11,538.2GPT-5.6
LiveBench
global average (0-100)
81.176GPT-5.6
SimpleBench
AVG@5 % (private set)
71.760.6GPT-5.6
Vals AI
Vals Index accuracy % · weighted finance + coding tasks
63.759.6GPT-5.6
Artificial Analysis
intelligence index · measured runs only
4738.2GPT-5.6

Raw scores are each board's own scale, so compare within a row, not across rows. Full table: the bench index.

GPT-5.6 at a glance

Developer
OpenAI
Status
Live since July 9, 2026
API price
$4 input / $20 output per 1M tokens · OpenAI, gpt-5.6
Context window
1.05M tokens

Live — Sol · Terra · Luna since Jul 9 · repriced Jul 30: Luna −80%, Terra −20%.

Three ChatGPT leaks in five hours were one release note — and the line none of them quoted is the real change

2026-08-06 · Three separate leaks landed in five hours claiming three separate ChatGPT changes: an upgraded GPT-5.6 Sol for Plus and Pro, a new reasoning slider rolling out, and unlimited chats coming for free users. All three are real.

OpenAI cut Luna 80% — Azure kept the old price, so Luna costs 5× more on Azure than direct

2026-08-04 · OpenAI cut GPT-5.6 Luna by 80% on July 30. We covered it. What we did not check — and should have, because it is the number that actually lands on a bill — is whether the clouds passed it on.

Full GPT-5.6 tracker →

Claude Sonnet 5 at a glance

Developer
Anthropic
Status
Live since June 30, 2026
API price
$2 input / $10 output per 1M tokens · Anthropic, claude-sonnet-5
Context window
1M tokens

Shipped Jun 30.

Full Claude Sonnet 5 tracker →

GPT-5.6 vs Claude Sonnet 5: quick answers

Which is better, GPT-5.6 or Claude Sonnet 5?

It depends on what you measure. GPT-5.6 scores higher than Claude Sonnet 5 on all 5 public leaderboards that rate both. GPT-5.6's widest lead is on LiveBench. On the ArtificialWatch Index, which takes the median of each model's percentile across the boards that score it, GPT-5.6 is at 70.2 and Claude Sonnet 5 at 36.7 (out of 100).

Which is newer, GPT-5.6 or Claude Sonnet 5?

GPT-5.6. It went live on July 9, 2026, 9 days after Claude Sonnet 5 (June 30, 2026).

Which is cheaper, GPT-5.6 or Claude Sonnet 5?

Claude Sonnet 5 is cheaper on both input and output. On each vendor's own API, GPT-5.6 is $4 input / $20 output per million tokens and Claude Sonnet 5 is $2 input / $10 output per million tokens. GPT-5.6's rate rises for prompts over 272K tokens. Prices as published in the models.dev catalog, as of 2026-09-25.

Which has the bigger context window, GPT-5.6 or Claude Sonnet 5?

GPT-5.6, at 1.05M tokens against Claude Sonnet 5's 1M, as each vendor lists it.

Who makes GPT-5.6 and Claude Sonnet 5?

GPT-5.6 is from OpenAI; Claude Sonnet 5 is from Anthropic.

Where do these numbers come from?

Each score is the leaderboard's own published figure, read by ArtificialWatch (table composed September 25, 2026). Boards measure different things, so raw scores are only comparable within a row.

More comparisons

GPT-6 Astra vs Claude Fable 5.1GPT-6 Astra vs GPT-5.6GPT-6 Astra vs Claude Opus 5Claude Fable 5.1 vs Claude Fable 5Claude Fable 5.1 vs Claude Opus 5Claude Opus 5 vs GPT-5.6Claude Opus 5 vs Claude Opus 4.8Claude Opus 5 vs Claude Fable 5Claude Opus 5 vs Kimi K3Kimi K3 vs Claude Fable 5Kimi K3 vs GLM-5.2Kimi K3 vs GPT-5.6DeepSeek V4 Pro vs GLM-5.2DeepSeek V4 Pro vs Kimi K3Grok 4.7 vs GPT-6 AstraGrok 4.7 vs Claude Opus 5Grok 4.7 vs Claude Fable 5.1Qwen 3.8 Max vs Claude Opus 5Qwen 3.8 Max vs Claude Fable 5Qwen 3.8 Max vs Kimi K3Qwen 3.8 Max vs DeepSeek V4 ProQwen 3.8 Max vs GLM-5.2Qwen 3.8 Max vs GPT-5.6GPT-6 Astra vs Claude Fable 5Claude Fable 5 vs Claude Opus 4.8Claude Fable 5 vs Claude Sonnet 5Claude Sonnet 5 vs Claude Opus 4.8Claude Sonnet 5 vs Claude Opus 5GLM-5.2 vs Claude Opus 4.8GLM-5.2 vs Claude Fable 5GLM-5.2 vs GPT-5.5Gemini 3.5 Flash vs Gemini 3.1 ProGemini 3.5 Flash vs Claude Sonnet 5GPT-6 Astra vs Claude Opus 4.8Claude Fable 5.1 vs Claude Opus 4.8Kimi K3 vs Claude Opus 4.8DeepSeek V4 Pro vs Claude Opus 4.8DeepSeek V4 Pro vs GPT-5.5DeepSeek V4 Pro vs Claude Fable 5GLM-5.3 vs Kimi K3GLM-5.3 vs GLM-5.2GLM-5.3 vs GPT-6 AstraGrok 4.6 vs Claude Opus 5Grok 4.6 vs Claude Fable 5Grok 4.6 vs GPT-5.6Grok 4.6 vs Grok 4.5Grok 4.6 vs Claude Sonnet 5Gemini 3.8 Flash vs Gemini 3.1 ProGemini 3.8 Flash vs Claude Opus 5Muse Spark 1.3 vs Gemini 3.8 FlashMuse Spark 1.3 vs GLM-5.3GPT-6 Sol vs GPT-6 AstraGPT-6 Sol vs Claude Opus 5GPT-6 Sol vs Claude Fable 5GPT-6 Luna vs GPT-6 AstraGPT-6 Sol vs GPT-6 LunaClaude Fable 5 vs GPT-5.6Claude Opus 4.8 vs GPT-5.5GPT-5.6 vs Claude Opus 4.8Claude Sonnet 5 vs GPT-5.5Claude Fable 5 vs GPT-5.5DeepSeek V4 Pro vs Claude Opus 5Qwen 3.8 Max vs GLM-5.3Qwen 3.8 Max vs Claude Opus 4.8GLM-5.3 Flash vs GLM-5.3GLM-5.3 Flash vs GLM-5.2MiniMax M3 vs Kimi K3MiniMax M3 vs DeepSeek V4 ProMiniMax M3 vs GLM-5.2MiniMax M3 vs GPT-5.5
sweeping every 60 seconds

Hear the minute the next one drops. Free.

A free Chrome push the moment a new model answers on a public API, an email about 15 minutes later. Texts or a phone call on paid plans.

  • Free forever
  • No card
  • Unsubscribe in one click

Also: the bench index · API pricing · the wire