← comparisons · head to head

DeepSeek V4 Pro vs Claude Opus 4.8

Of the 5 public leaderboards that rate both, DeepSeek V4 Pro scores higher on 2 and Claude Opus 4.8 on 3. On the ArtificialWatch Index, which takes the median of each model's percentile across the boards that score it, DeepSeek V4 Pro is at 37.1 and Claude Opus 4.8 at 47.8 (out of 100).

Scores as published by each board; table composed September 24, 2026.

Get the next launch alert, free →

Leaderboard by leaderboard

BoardDeepSeek V4 ProClaude Opus 4.8Higher
LMArena
Elo (community votes)
1,580.51,555.2DeepSeek V4 Pro
LiveBench
global average (0-100)
77.476.2DeepSeek V4 Pro
SimpleBench
AVG@5 % (private set)
50.964.8Claude Opus 4.8
Vals AI
Vals Index accuracy % · weighted finance + coding tasks
52.460.9Claude Opus 4.8
Artificial Analysis
intelligence index · measured runs only
3641.8Claude Opus 4.8

Raw scores are each board's own scale, so compare within a row, not across rows. Full table: the bench index.

DeepSeek V4 Pro at a glance

Developer
DeepSeek
Status
Live since April 24, 2026
API price
$0.435 input / $0.87 output per 1M tokens · DeepSeek, deepseek-v4-pro
Context window
1M tokens

Live · open weights · superseded by V4.1 Flash — per DeepSeek's own changelog rather than our own reading, deepseek-v4-flash routes to V4.1 Flash; the matching deepseek-v4-pro cutover set for 04:00 UTC Sep 14 was reversed on its eve and V4 Pro is still served at unchanged billing, which is search-grade too — but our own bench refresh still scored deepseek-v4-pro-0813 as a separate model on Sep 16 · our sweep read the bare deepseek-flash alias on Sep 10 at $0.15/$0.60 per 1M, 1M ctx.

DeepSeek V4-Pro lands at $0.435/$0.87 — and "model names remain unchanged" hides a 2.7x price fork

2026-08-13 · DeepSeek launched V4-Pro today on app, web and API. The headline features are agent upgrades, selectable reasoning effort across V4-Pro and V4-Flash, and native OpenAI Responses API support optimised for Codex. The line worth reading twice is the quiet one: "Model names remain unchanged."

Full DeepSeek V4 Pro tracker →

Claude Opus 4.8 at a glance

Developer
Anthropic
Status
Live since May 27, 2026
API price
$5 input / $25 output per 1M tokens · Anthropic, claude-opus-4-8
Context window
1M tokens

Live.

Full Claude Opus 4.8 tracker →

DeepSeek V4 Pro vs Claude Opus 4.8: quick answers

Which is better, DeepSeek V4 Pro or Claude Opus 4.8?

It depends on what you measure. Of the 5 public leaderboards that rate both, DeepSeek V4 Pro scores higher on 2 and Claude Opus 4.8 on 3. DeepSeek V4 Pro's widest lead is on LiveBench; Claude Opus 4.8's widest lead is on Vals AI. On the ArtificialWatch Index, which takes the median of each model's percentile across the boards that score it, DeepSeek V4 Pro is at 37.1 and Claude Opus 4.8 at 47.8 (out of 100).

Which is newer, DeepSeek V4 Pro or Claude Opus 4.8?

Claude Opus 4.8. It went live on May 27, 2026, 33 days after DeepSeek V4 Pro (April 24, 2026).

Which is cheaper, DeepSeek V4 Pro or Claude Opus 4.8?

DeepSeek V4 Pro is cheaper on both input and output. On each vendor's own API, DeepSeek V4 Pro is $0.435 input / $0.87 output per million tokens and Claude Opus 4.8 is $5 input / $25 output per million tokens. Prices as published in the models.dev catalog, as of 2026-09-24.

Which has the bigger context window, DeepSeek V4 Pro or Claude Opus 4.8?

Neither. Both take 1M tokens.

Who makes DeepSeek V4 Pro and Claude Opus 4.8?

DeepSeek V4 Pro is from DeepSeek; Claude Opus 4.8 is from Anthropic.

Where do these numbers come from?

Each score is the leaderboard's own published figure, read by ArtificialWatch (table composed September 24, 2026). Boards measure different things, so raw scores are only comparable within a row.

More comparisons

GPT-6 Astra vs Claude Fable 5.1GPT-6 Astra vs GPT-5.6GPT-6 Astra vs Claude Opus 5Claude Fable 5.1 vs Claude Fable 5Claude Fable 5.1 vs Claude Opus 5Claude Opus 5 vs GPT-5.6Claude Opus 5 vs Claude Opus 4.8Claude Opus 5 vs Claude Fable 5Claude Opus 5 vs Kimi K3Kimi K3 vs Claude Fable 5Kimi K3 vs GLM-5.2Kimi K3 vs GPT-5.6DeepSeek V4 Pro vs GLM-5.2DeepSeek V4 Pro vs Kimi K3Grok 4.7 vs GPT-6 AstraGrok 4.7 vs Claude Opus 5Grok 4.7 vs Claude Fable 5.1Qwen 3.8 Max vs Claude Opus 5Qwen 3.8 Max vs Claude Fable 5Qwen 3.8 Max vs Kimi K3Qwen 3.8 Max vs DeepSeek V4 ProQwen 3.8 Max vs GLM-5.2Qwen 3.8 Max vs GPT-5.6GPT-6 Astra vs Claude Fable 5Claude Fable 5 vs Claude Opus 4.8Claude Fable 5 vs Claude Sonnet 5Claude Sonnet 5 vs Claude Opus 4.8Claude Sonnet 5 vs Claude Opus 5GLM-5.2 vs Claude Opus 4.8GLM-5.2 vs Claude Fable 5GLM-5.2 vs GPT-5.5Gemini 3.5 Flash vs Gemini 3.1 ProGemini 3.5 Flash vs Claude Sonnet 5GPT-6 Astra vs Claude Opus 4.8Claude Fable 5.1 vs Claude Opus 4.8Kimi K3 vs Claude Opus 4.8DeepSeek V4 Pro vs GPT-5.5DeepSeek V4 Pro vs Claude Fable 5
Hear the minute the next one drops.

Watch free: Chrome push the moment a new model answers on a public API, an email about 15 minutes later. Also: the bench index · API pricing · the wire