← Model explorer

Sourced model profile · Current through July 21, 2026

DeepSeek-V4-Pro

An unrestricted open-weight DeepSeek language model documented for generation and question answering.

Developed byDeepSeek
Released
April 24, 2026
LMArena rating
1457.5
Snapshot date
July 16, 2026

What the evidence says

These facts come from the reviewed Epoch AI catalog artifact and its linked primary record. They describe the model. They do not rank it.

Developer [1]
DeepSeek
Released [1]
April 24, 2026
Domains [1]
Language
Tasks [1]
Language modeling/generation · Question answering
Access [1]
Open weights (unrestricted)
Weights [1]
Open

Where it may fit

These are evidence-bounded screening suggestions, not performance claims or purchase recommendations.

Self-hosted language evaluation

Consider it when open weights and unrestricted access are requirements for a language-model evaluation.

Evidence fields: domains, access, openWeights [1]

Question-answering prototypes

Consider it for question-answering research that benefits from inspectable, deployable weights.

Evidence fields: tasks, access, openWeights [1]

Limitations

The record does not provide hardware requirements, serving cost, context window, latency, safety evaluations, or a directly matched Arena score. Deployment suitability still requires independent testing.

Current standing

According to the LMArena leaderboard dataset published July 16, 2026, DeepSeek-V4-Pro holds a rating of 1457.5 in the style-controlled text category, with a reported 95% interval of 1453.1 to 1462.0 across 43,062 battles. This is one source's point-in-time observation, not a composite score or a claim of overall superiority.

Head-to-head comparisons

Source-backed pair pages that put DeepSeek-V4-Pro next to another documented model. None declares a winner beyond what a named source's numbers say.

Across the boards

Each card below is one named authority's own current view of DeepSeek-V4-Pro. Model Gauntlet mirrors these sources with attribution and does not combine them into a single score or rank DeepSeek-V4-Pro across them.

LMArena overall

Human preference, style-controlled text

LMArena rating
1457.1
Board rank
#43
Reported interval
1452.7 to 1461.5
Battles
44,418

Published July 20, 2026. LMArena leaderboard dataset, CC-BY-4.0.

Overall leaderboard →

LMArena coding arena

Human preference, coding arena

LMArena rating
1501.4
Board rank
#48
Reported interval
1494.8 to 1508.1
Battles
13,142

Published July 20, 2026. LMArena leaderboard dataset, CC-BY-4.0.

Coding arena board →

LMArena creative writing arena

Human preference, creative writing arena

LMArena rating
1443.6
Board rank
#36
Reported interval
1435.3 to 1451.8
Battles
7,179

Published July 20, 2026. LMArena leaderboard dataset, CC-BY-4.0.

Creative writing arena board →

LMArena instruction following arena

Human preference, instruction following arena

LMArena rating
1453.5
Board rank
#38
Reported interval
1447.2 to 1459.8
Battles
14,929

Published July 20, 2026. LMArena leaderboard dataset, CC-BY-4.0.

Instruction following arena board →

LMArena hard prompts arena

Human preference, hard prompts arena

LMArena rating
1480.2
Board rank
#41
Reported interval
1475.1 to 1485.4
Battles
29,323

Published July 20, 2026. LMArena leaderboard dataset, CC-BY-4.0.

Hard prompts arena board →

LMArena multi-turn arena

Human preference, multi-turn arena

LMArena rating
1472.8
Board rank
#37
Reported interval
1465.0 to 1480.6
Battles
8,027

Published July 20, 2026. LMArena leaderboard dataset, CC-BY-4.0.

Multi-turn arena board →

LMArena long queries arena

Human preference, long queries arena

LMArena rating
1473.2
Board rank
#37
Reported interval
1467.1 to 1479.2
Battles
19,139

Published July 20, 2026. LMArena leaderboard dataset, CC-BY-4.0.

Long queries arena board →

LMArena math arena

Human preference, math arena

LMArena rating
1444.2
Board rank
#56
Reported interval
1431.7 to 1456.7
Battles
2,392

Published July 20, 2026. LMArena leaderboard dataset, CC-BY-4.0.

Math arena board →

EQ-Bench 3

Emotional intelligence

Judge Elo
1329.7
Source rank
#5
Source label
deepseek/deepseek-v4-pro

EQ-Bench 3 is a single-source subjective judge-based pairwise Elo for emotional-intelligence roleplays, appropriate only for that lane.

Source last updated May 10, 2026. EQ-Bench 3 canonical Elo leaderboard, EQ-bench/eqbench3, MIT.

EQ-Bench 3 source ↗