← Overall rankings

LMArena Multi-turn arena · published July 20, 2026

Anthropic leads the LMArena multi-turn arena.

As of July 20, 2026, Anthropic leads the LMArena multi-turn arena with claude-opus-4-7 at 1518.8 across 9,076 battles, according to LMArena.

Current through July 20, 2026. 145 published snapshots. 377 of 381 models resolved (4 held for exact-match review).

LMArena multi-turn arena rating, published July 20, 2026. Top 25 of 322 rows.
RankLMArena ratingOrganizationModelReported intervalGap to leaderBattles
011518.8AnthropicClaude Opus 4.71511.1 to 1526.4Leader9,076
021517.9Anthropicclaude-opus-4-6-thinking1510.9 to 1524.8-0.910,654
031517.5Anthropicclaude-opus-4-7-thinking1509.8 to 1525.2-1.38,703
041514.9AnthropicClaude Fable 51499.7 to 1530.1-3.91,637
051510.6Anthropicclaude-opus-4-61503.9 to 1517.3-8.211,653
061506.3Anthropicclaude-opus-4-8-thinking1497.2 to 1515.4-12.55,367
071505.2Metamuse-spark-1.11487.2 to 1523.2-13.61,199
081499.4Anthropicclaude-opus-4-81490.4 to 1508.4-19.45,652
091495.3GoogleGemini 3 Pro1487.5 to 1503.2-23.56,766
101494.4OpenAIgpt-5.2-chat-latest-202602101486.4 to 1502.5-24.46,487
111493.2OpenAIgpt-5.4-high1486.3 to 1500.2-25.610,720
121492.5Googlegemini-3.1-pro-preview1486.2 to 1498.8-26.314,472
131492.2MetaMuse Spark1478.7 to 1505.7-26.62,164
141489.2OpenAIgpt-5.5-high1481.1 to 1497.2-29.67,728
151488.6Googlegemini-3.5-flash-high1474.1 to 1503.2-30.21,843
161487.3Anthropicclaude-opus-4-5-20251101-thinking-32k1479.4 to 1495.1-31.56,468
171484.0Anthropicclaude-opus-4-5-202511011477.8 to 1490.2-34.812,676
181483.6Anthropicclaude-sonnet-4-61476.5 to 1490.6-35.210,182
191483.5Googlegemini-3-flash1474.8 to 1492.1-35.35,320
201483.4Z.aiglm-5.11474.2 to 1492.6-35.44,945
211483.4Alibabaqwen3.7-max-preview1459.8 to 1506.9-35.4659
221482.7Googlegemini-3.5-flash-medium1469.8 to 1495.6-36.12,317
231482.4OpenAIgpt-5.5-instant1473.0 to 1491.9-36.44,906
241482.0xAIgrok-4.20-beta11472.4 to 1491.6-36.84,384
251481.8OpenAIgpt-5.51473.8 to 1489.7-37.08,151

Frontier race

Category Kings: Multi-turn arena frontier by organization

Anthropic holds the multi-turn arena frontier at 1519 as of July 20, 2026.

Multi-turn arena frontier ratings by organization over timeAnthropic holds the multi-turn arena frontier at 1519 as of July 20, 2026. 9 organizations are tracked from January 5, 2025 to July 20, 2026.12201310140014902025-012025-052025-092026-012026-042026-07Anthropic 1519Meta 1505Google 1495OpenAI 1494Alibaba 1483Z.ai 1483xAI 1482DeepSeek 1473Mistral AI 1432

Method: Frontier is the maximum resolved LMArena rating per organization on each publish date. Organizations limited to the union of the seven tracked homepage providers and organizations in the latest board top 25 by officialRank. Quarantined dates are excluded, and this category series keeps its own method epoch and is never joined to the legacy archive.

Multi-turn arena frontier by organization: latest resolved score per organization, from the reviewed LMArena category artifact.
OrganizationLatest modelLatest LMArena ratingFirst tracked date
Anthropicanthropic/claude-opus-4-71518.8January 5, 2025
Metameta/muse-spark-1.11505.2January 5, 2025
Googlegoogle/gemini-3-pro1495.3January 5, 2025
OpenAIopenai/gpt-5.2-chat-latest-202602101494.4January 5, 2025
Alibabaalibaba/qwen3.7-max-preview1483.4January 5, 2025
Z.aizai/glm-5.11483.4January 5, 2025
xAIxai/grok-4.20-beta11482.0January 5, 2025
DeepSeekdeepseek/deepseek-v4-pro1472.8January 5, 2025
Mistral AImistral/mistral-medium-3.51431.8January 5, 2025
Anthropic leads the multi-turn arena frontier at 1518.8 on July 20, 2026. Frontier is the highest resolved score per organization per publish date. Quarantined dates are excluded and this category is never joined to the legacy archive.

Definitions

What these terms mean.

Multi-turn arena
The LMArena style-controlled battles with more than one conversational turn. It reflects preference across a dialogue, not long-horizon task completion.
Frontier by organization
The highest resolved score for each organization on each publish date. It tracks how far each lab has pushed this category, not an average of all its models.
Confidence interval
The reported 95% range LMArena publishes with each rating. Overlapping intervals mean this snapshot does not separate those ranks decisively.

Multi-turn arena questions

Which model leads the Multi-turn arena right now?

According to the LMArena multi-turn conversation board published July 20, 2026, Anthropic leads with claude-opus-4-7 at 1518.8, across 9,076 battles. That is 0.9 points ahead of the next board entry inside this one snapshot.

Is this a Model Gauntlet multi-turn conversation ranking?

No. This board mirrors one named source, the LMArena multi-turn conversation arena, published July 20, 2026. Model Gauntlet publishes no composite or consensus score and runs no evaluation of its own.

How many models and organizations are in the multi-turn conversation board?

The reviewed artifact covers 381 models across 61 organizations, with 377 models resolved to exact identities and 4 held for review. The history spans 145 snapshots from January 5, 2025 to July 20, 2026.

Why do some scores have overlapping confidence intervals?

LMArena publishes a 95% confidence interval with each rating. Where two intervals overlap, this single snapshot does not separate those ranks decisively, so treat small multi-turn conversation gaps with caution.