LMArena snapshot · published July 16, 2026

Anthropic leads this LMArena snapshot.

This leaderboard ranks every resolved model in one published LMArena rating for text models, using LMArena's official ranks. A second table summarizes the leading model from each of seven organizations. Confidence intervals and battle counts appear alongside each score.

Model rankings · LMArena overall

Every resolved model, ranked.

LMArena official rank and rating, style-controlled text category, published July 20, 2026. Top 25 of 324 resolved rows.
RankLMArena ratingModelOrganizationReported intervalGap to leaderBattles
011506.8Claude Fable 5Anthropic1499.7 to 1513.8Leader9,798
021504.2Claude Opus 4.6 ThinkingAnthropic1500.4 to 1507.9-2.662,355
031501.9Claude Opus 4.7 ThinkingAnthropic1497.7 to 1506.1-4.949,798
041497.9Claude Opus 4.6Anthropic1494.3 to 1501.6-8.966,131
051495.1Muse Spark 1.1Meta1487.6 to 1502.6-11.77,074
061493.7Claude Opus 4.7Anthropic1489.5 to 1497.9-13.150,890
071487.4Muse SparkMeta1481.5 to 1493.2-19.413,562
091486.4GPT-5.6 Sol X-HighOpenAI1478.0 to 1494.7-20.45,467
101485.8Gemini 3 ProGoogle1481.9 to 1489.6-21.041,279
111485.6Gemini 3.1 Pro PreviewGoogle1482.1 to 1489.1-21.283,386
121484.6Claude Opus 4.8 ThinkingAnthropic1479.5 to 1489.8-22.230,024
131481.0GPT-5.5 HighOpenAI1476.6 to 1485.4-25.844,927
141477.9GPT-5.4 HighOpenAI1473.9 to 1481.8-28.958,116
151476.3Gemini 3.5 Flash HighGoogle1469.8 to 1482.9-30.510,094
161475.7GPT-5.2 Chat Latest (2026-02-10)OpenAI1471.6 to 1479.9-31.134,416
171475.7GPT-5.5OpenAI1471.3 to 1480.1-31.146,328
181475.3Qwen3.7 Max PreviewAlibaba1465.3 to 1485.4-31.53,715
191474.8Gemini 3.5 Flash MediumGoogle1468.7 to 1481.0-32.013,158
201474.4Claude Opus 4.8Anthropic1469.3 to 1479.6-32.430,670
211474.2Grok 4.20 Beta 1xAI1469.5 to 1478.8-32.626,827
221473.3Grok 4.20 Beta 0309 ReasoningxAI1469.4 to 1477.1-33.559,422
231473.2GPT-5.5 InstantOpenAI1468.1 to 1478.3-33.626,015
241473.0Gemini 3 FlashGoogle1468.6 to 1477.4-33.830,695
251473.0Claude Opus 4.5 Thinking 32K (2025-11-01)Anthropic1469.1 to 1476.9-33.837,049
261472.5Claude Sonnet 4.6Anthropic1468.6 to 1476.3-34.356,335
Show the remaining 299 resolved rows
Rows 26 and beyond of the LMArena model rankings published July 20, 2026.
RankLMArena ratingModelOrganizationReported intervalGap to leaderBattles
271470.4Grok 4.20 Multi Agent Beta 0309xAI1466.5 to 1474.3-36.458,221
281469.9GLM-5.1Z.ai1465.3 to 1474.5-36.929,919
291469.5GLM 5.2 (Max)Z.ai1463.6 to 1475.5-37.317,103
301469.2Claude Opus 4.5 (2025-11-01)Anthropic1466.0 to 1472.4-37.670,973
311467.9Ernie 5.1Baidu1463.2 to 1472.5-38.937,445
321466.7GPT 5.4OpenAI1462.8 to 1470.6-40.161,064
331465.9Mimo V2.5 ProXiaomi1461.4 to 1470.4-40.941,409
341465.9Grok 4.1 ThinkingxAI1462.7 to 1469.1-40.965,461
351465.7Grok 4.5xAI1458.1 to 1473.3-41.16,942
361464.8Qwen3.5 Max PreviewAlibaba1459.8 to 1469.8-42.021,479
371461.4Claude Sonnet 5 HighAnthropic1455.1 to 1467.8-45.412,645
381461.2Kimi K2.6Moonshot AI1456.6 to 1465.7-45.637,711
391460.8Qwen3.7 PlusAlibaba1455.1 to 1466.4-46.021,513
401460.3Qwen3.6 Max PreviewAlibaba1451.9 to 1468.7-46.55,191
411459.5Grok 4.1xAI1456.2 to 1462.7-47.367,587
421458.8Gemini 3 Flash (thinking Minimal)Google1455.6 to 1461.9-48.083,526
431457.1DeepSeek V4 ProDeepSeek1452.7 to 1461.5-49.744,418
441456.7Glm 5Z.ai1452.3 to 1461.0-50.127,789
451456.0DeepSeek V4 Pro ThinkingDeepSeek1451.5 to 1460.5-50.842,294
461456.0Dola Seed 2.0 ProByteDance1452.3 to 1459.6-50.867,083
471455.5Claude Sonnet 4 5 20250929 Thinking 32kAnthropic1452.7 to 1458.3-51.382,325
481455.1Claude Sonnet 4 5 20250929Anthropic1452.2 to 1458.0-51.780,727
491454.7GPT 5.1 HighOpenAI1450.9 to 1458.4-52.140,779
501450.9Gemma 4 31bGoogle1443.3 to 1458.5-55.95,879
511449.3GPT 5.4 Mini HighOpenAI1445.4 to 1453.3-57.556,941
521449.2Kimi K2.5 ThinkingMoonshot AI1445.6 to 1452.7-57.661,978
531449.1Claude Opus 4 1 20250805 Thinking 16kAnthropic1445.7 to 1452.6-57.749,757
541448.8GPT 5.3 Chat LatestOpenAI1444.5 to 1453.2-58.032,982
551448.8Ernie 5.0 Preview 1203Baidu1442.3 to 1455.3-58.09,734
561448.3Mimo V2 ProXiaomi1443.6 to 1453.1-58.524,478
571447.0Claude Opus 4 1 20250805Anthropic1444.1 to 1450.0-59.877,243
581446.8Ernie 5.0 0110Baidu1442.9 to 1450.8-60.035,225
601445.6Gemini 2.5 ProGoogle1443.1 to 1448.1-61.2124,385
611445.0Minimax M3MiniMax1439.7 to 1450.4-61.827,238
621444.7GPT 4.5 Preview 2025 02 27OpenAI1439.0 to 1450.4-62.114,547
631443.5Qwen3.6 PlusAlibaba1439.2 to 1447.8-63.343,163
641443.1Chatgpt 4o Latest 20250326OpenAI1440.3 to 1445.9-63.782,388
651442.7Grok 4.3xAI1438.3 to 1447.0-64.145,525
661442.2Qwen3.5 397b A17bAlibaba1438.5 to 1446.0-64.657,410
671442.1Glm 4.7Z.ai1436.0 to 1448.2-64.712,094
681438.6GPT 5.1OpenAI1435.0 to 1442.3-68.243,398
691438.3DeepSeek V4 Flash ThinkingDeepSeek1433.9 to 1442.7-68.544,034
701438.2Gemma 4 26b A4bGoogle1430.6 to 1445.9-68.65,798
711437.4GPT 5.2 HighOpenAI1433.7 to 1441.0-69.447,939
721436.5Glm 5v TurboZ.ai1428.8 to 1444.1-70.36,753
731436.1DeepSeek V4 FlashDeepSeek1431.8 to 1440.5-70.744,220
741435.8Longcat Flash Chat 2602 ExpMeituan1431.1 to 1440.4-71.028,080
751434.8Qwen3 Max PreviewAlibaba1430.3 to 1439.3-72.027,696
761434.5GPT 5.2OpenAI1431.2 to 1437.8-72.376,740
771433.8GPT 5 HighOpenAI1429.3 to 1438.3-73.031,889
781433.1MiMo V2.5Xiaomi1428.6 to 1437.5-73.742,281
791431.8Gemini 3.1 Flash Lite PreviewGoogle1428.0 to 1435.5-75.060,792
801431.6Kimi K2.5 InstantMoonshot AI1425.1 to 1438.2-75.28,177
811431.0Grok 4 1 Fast ReasoningxAI1427.7 to 1434.2-75.856,767
821430.9O3 2025 04 16OpenAI1427.3 to 1434.5-75.959,695
831430.1Mimo V2 OmniXiaomi1424.3 to 1435.8-76.719,464
841429.6Kimi K2 Thinking TurboMoonshot AI1426.5 to 1432.8-77.261,948
851427.6Mistral Medium 3.5Mistral AI1421.0 to 1434.2-79.211,011
861426.9Amazon Nova Experimental Chat 26 02 10Amazon1417.1 to 1436.8-79.93,419
871426.7GPT 5 ChatOpenAI1422.4 to 1431.0-80.131,519
881425.2Glm 4.6Z.ai1421.3 to 1429.1-81.635,613
891424.9DeepSeek V3.2DeepSeek1421.3 to 1428.5-81.947,211
901424.7DeepSeek V3.2 Exp ThinkingDeepSeek1418.1 to 1431.2-82.19,067
911424.5Claude Opus 4 20250514 Thinking 16kAnthropic1420.1 to 1428.8-82.336,860
921424.5Nvidia Nemotron 3 Ultra 550b A55b Nvfp4NVIDIA1417.4 to 1431.6-82.310,283
931424.1Qwen3 Max 2025 09 23Alibaba1417.6 to 1430.5-82.79,148
941423.1Qwen3 235b A22b Instruct 2507Alibaba1420.5 to 1425.7-83.797,091
951422.9DeepSeek V3.2 ThinkingDeepSeek1419.3 to 1426.6-83.941,026
961422.7DeepSeek V3.2 ExpDeepSeek1416.3 to 1429.1-84.111,914
971422.1DeepSeek R1 0528DeepSeek1416.4 to 1427.7-84.718,452
981420.8Grok 4 Fast ChatxAI1413.1 to 1428.4-86.06,807
991418.4Ernie 5.0 Preview 1022Baidu1409.6 to 1427.2-88.44,702
1001417.9Kimi K2 0905 PreviewMoonshot AI1411.4 to 1424.4-88.911,771
1011417.8Minimax M2.7MiniMax1413.6 to 1422.0-89.049,135
1021417.5Kimi K2 0711 PreviewMoonshot AI1412.6 to 1422.4-89.327,605
1031417.5DeepSeek V3.1DeepSeek1411.5 to 1423.5-89.314,941
1041417.4DeepSeek V3.1 Terminus ThinkingDeepSeek1407.4 to 1427.4-89.43,456
1051417.3Qwen3.5 122b A10bAlibaba1412.9 to 1421.7-89.528,501
1061417.0DeepSeek V3.1 ThinkingDeepSeek1410.4 to 1423.6-89.811,723
1071415.4Amazon Nova Experimental Chat 26 01 10Amazon1405.5 to 1425.3-91.43,407
1081415.3Mistral Large 3Mistral AI1411.9 to 1418.7-91.550,153
1091415.3Qwen3 Vl 235b A22b InstructAlibaba1408.8 to 1421.8-91.511,498
1101415.3DeepSeek V3.1 TerminusDeepSeek1405.6 to 1424.9-91.53,690
1111413.7GPT 4.1 2025 04 14OpenAI1410.0 to 1417.4-93.150,929
1121412.5Claude Opus 4 20250514Anthropic1408.2 to 1416.8-94.344,168
1131412.2Claude Haiku 4 5 20251001Anthropic1409.5 to 1414.9-94.6107,630
1141412.0Hunyuan Hy3 PreviewTencent1404.5 to 1419.6-94.86,637
1151411.5Grok 3 Preview 02 24xAI1407.2 to 1415.9-95.332,894
1161410.9Glm 4.5Z.ai1406.0 to 1415.8-95.924,285
1171410.3Gemini 2.5 FlashGoogle1407.8 to 1412.7-96.5124,312
1181409.6Grok 4 0709xAI1405.7 to 1413.5-97.241,345
1191409.5Mistral Medium 2508Mistral AI1406.9 to 1412.2-97.393,809
1201408.7Qwen3.5 27bAlibaba1404.3 to 1413.2-98.127,308
1211404.2Gemini 2.5 Flash Preview 09 2025Google1400.2 to 1408.3-102.632,881
1221404.0Grok 4 Fast ReasoningxAI1399.0 to 1409.0-102.818,705
1231403.4GPT 5.4 Nano HighOpenAI1399.4 to 1407.3-103.455,860
1241403.0Qwen3 235b A22b No ThinkingAlibaba1398.5 to 1407.5-103.838,175
1251402.0O1 2024 12 17OpenAI1397.6 to 1406.5-104.827,807
1261401.3Qwen3 Next 80b A3b InstructAlibaba1396.5 to 1406.1-105.522,853
1271401.2Longcat Flash ChatMeituan1394.9 to 1407.6-105.611,384
1281399.2Claude Sonnet 4 20250514 Thinking 32kAnthropic1394.8 to 1403.6-107.635,065
1291399.0Qwen3 235b A22b Thinking 2507Alibaba1392.5 to 1405.6-107.88,985
1301398.1DeepSeek R1DeepSeek1393.2 to 1403.0-108.718,524
1311397.3Qwen3.5 FlashAlibaba1393.4 to 1401.1-109.555,880
1321395.7Qwen3.5 35b A3bAlibaba1391.4 to 1400.1-111.129,157
1331395.5DeepSeek V3 0324DeepSeek1391.6 to 1399.4-111.345,480
1341395.4Qwen3 Vl 235b A22b ThinkingAlibaba1388.6 to 1402.2-111.47,937
1351395.2Hunyuan Vision 1.5 ThinkingTencent1382.9 to 1407.4-111.62,217
1361394.9Step 3.5 FlashStepFun1391.2 to 1398.6-111.955,180
1371394.3Amazon Nova Experimental Chat 12 10Amazon1384.7 to 1403.8-112.53,677
1381392.8Mimo V2 Flash (non Thinking)Xiaomi1389.2 to 1396.4-114.046,546
1391390.6Minimax M2.5MiniMax1386.6 to 1394.6-116.241,092
1401390.1GPT 5 Mini HighOpenAI1385.4 to 1394.7-116.727,001
1411389.9O4 Mini 2025 04 16OpenAI1386.0 to 1393.9-116.945,414
1421389.2Claude Sonnet 4 20250514Anthropic1384.8 to 1393.5-117.640,277
1431388.3O1 PreviewOpenAI1383.3 to 1393.3-118.531,122
1441387.5Qwen3 Coder 480b A35b InstructAlibaba1382.6 to 1392.5-119.325,694
1451387.3Claude 3 7 Sonnet 20250219 Thinking 32kAnthropic1383.1 to 1391.5-119.538,814
1461387.1Mimo V2 Flash (thinking)Xiaomi1380.9 to 1393.3-119.710,937
1471386.9Hunyuan T1 20250711Tencent1378.2 to 1395.5-119.94,698
1481386.8Mistral Medium 2505Mistral AI1382.1 to 1391.5-120.033,193
1491384.2Minimax M2.1 PreviewMiniMax1379.0 to 1389.4-122.617,072
1501383.1Qwen3 30b A3b Instruct 2507Alibaba1378.2 to 1388.0-123.723,715
1511382.8GPT 4.1 Mini 2025 04 14OpenAI1378.5 to 1387.1-124.039,298
1521382.2Hunyuan Turbos 20250416Tencent1375.8 to 1388.7-124.610,726
1531379.6Gemini 2.5 Flash Lite Preview 09 2025 No ThinkingGoogle1376.2 to 1383.1-127.247,176
1551377.2Glm 4.6vZ.ai1365.9 to 1388.5-129.62,801
1561374.8Qwen3 235b A22bAlibaba1370.1 to 1379.5-132.026,257
1571374.3Gemini 2.5 Flash Lite Preview 06 17 ThinkingGoogle1369.8 to 1378.9-132.532,864
1581374.2Qwen2.5 MaxAlibaba1370.1 to 1378.3-132.632,613
1591373.1Glm 4.5 AirZ.ai1368.8 to 1377.3-133.731,062
1601373.1Claude 3 5 Sonnet 20241022Anthropic1369.9 to 1376.2-133.788,322
1611371.4Claude 3 7 Sonnet 20250219Anthropic1367.4 to 1375.3-135.443,161
1621369.5Qwen3 Next 80b A3b ThinkingAlibaba1363.7 to 1375.4-137.313,678
1641367.9Glm 4.7 FlashZ.ai1362.1 to 1373.7-138.911,706
1651366.2Amazon Nova Experimental Chat 11 10Amazon1361.9 to 1370.6-140.625,322
1661365.7Gemma 3 27b ItGoogle1362.0 to 1369.4-141.147,500
1671363.8Minimax M1MiniMax1359.5 to 1368.0-143.035,156
1681363.5O3 Mini HighOpenAI1358.2 to 1368.7-143.318,589
1691362.0Grok 3 Mini HighxAI1356.6 to 1367.3-144.816,949
1701361.6Nvidia Nemotron 3 Super 120b A12bNVIDIA1354.3 to 1368.8-145.27,542
1711360.3Gemini 2.0 Flash 001Google1356.5 to 1364.1-146.543,741
1721358.5DeepSeek V3DeepSeek1353.7 to 1363.2-148.321,770
1731357.5Mistral Small 2506Mistral AI1352.3 to 1362.7-149.317,696
1741356.7Grok 3 Mini BetaxAI1351.7 to 1361.7-150.122,688
1761353.9Command A 03 2025Cohere1350.5 to 1357.4-152.956,224
1771353.5Glm 4.5vZ.ai1345.2 to 1361.9-153.34,955
1781353.4Gemini 2.0 Flash Lite Preview 02 05Google1349.1 to 1357.7-153.424,955
1791352.5GPT Oss 120bOpenAI1348.2 to 1356.9-154.330,611
1801351.0Gemini 1.5 Pro 002Google1347.7 to 1354.4-155.855,606
1811349.7Amazon Nova Experimental Chat 10 20Amazon1343.5 to 1355.9-157.111,460
1821348.9Hunyuan Turbos 20250226Tencent1337.1 to 1360.6-157.92,220
1831348.3Step 3StepFun1340.9 to 1355.8-158.56,533
1841347.7O3 MiniOpenAI1344.2 to 1351.2-159.157,313
1851347.7Amazon Nova Experimental Chat 10 09Amazon1336.8 to 1358.5-159.12,824
1861347.5Llama 3.1 Nemotron Ultra 253b V1NVIDIA1335.8 to 1359.1-159.32,549
1871347.2Qwen3 32bAlibaba1337.7 to 1356.7-159.63,926
1881346.7Mercury 2Inception AI1336.1 to 1357.3-160.13,118
1891346.2Ling Flash 2.0Ant Group1339.0 to 1353.5-160.66,998
1901346.1Qwen Plus 0125Alibaba1337.8 to 1354.4-160.75,819
1911346.0Minimax M2MiniMax1338.3 to 1353.8-160.86,860
1921345.8GPT 4o 2024 05 13OpenAI1342.3 to 1349.2-161.0112,881
1931343.2Nvidia Llama 3.3 Nemotron Super 49b V1.5NVIDIA1333.2 to 1353.1-163.63,344
1941342.8Glm 4 Plus 0111Z.ai1334.4 to 1351.3-164.05,760
1951342.5Claude 3 5 Sonnet 20240620Anthropic1339.0 to 1345.9-164.382,419
1961341.9Gemma 3 12b ItGoogle1332.4 to 1351.4-164.93,829
1971340.7Hunyuan Turbo 0110Tencent1329.1 to 1352.2-166.12,290
1981337.2GPT 5 Nano HighOpenAI1330.3 to 1344.1-169.68,260
1991337.0O1 MiniOpenAI1333.4 to 1340.7-169.851,981
2001336.8Nova 2 LiteAmazon1330.6 to 1342.9-170.012,211
2011336.1Qwq 32bAlibaba1331.7 to 1340.5-170.725,376
2021335.6Grok 2 2024 08 13xAI1331.9 to 1339.3-171.263,498
2031335.3Gemini Advanced 0514Google1330.2 to 1340.5-171.550,148
2041335.1GPT 4o 2024 08 06OpenAI1330.9 to 1339.3-171.745,499
2051334.8Llama 3.1 405b Instruct Bf16Meta1331.1 to 1338.5-172.041,375
2061333.9Step 2 16k Exp 202412StepFun1325.3 to 1342.5-172.94,833
2071333.0Llama 3.1 405b Instruct Fp8Meta1329.4 to 1336.6-173.859,656
2081330.0Olmo 3.1 32b InstructAi21323.9 to 1336.1-176.812,205
2091328.8Molmo 2 8bAi21307.4 to 1350.1-178.0798
2111328.0Llama 3.3 Nemotron 49b Super V1NVIDIA1315.8 to 1340.1-178.82,218
2121327.2Qwen3 30b A3bAlibaba1322.4 to 1331.9-179.626,469
2131327.0Llama 4 Maverick 17b 128e InstructMeta1322.8 to 1331.3-179.839,951
2141326.2Hunyuan Large 2025 02 10Tencent1316.5 to 1335.9-180.63,738
2151324.0GPT 4 Turbo 2024 04 09OpenAI1320.1 to 1327.9-182.898,114
2161323.7Claude 3 5 Haiku 20241022Anthropic1320.5 to 1327.0-183.169,934
2171323.6Gemini 1.5 Pro 001Google1319.6 to 1327.6-183.279,138
2181323.4DeepSeek V2.5 1210DeepSeek1315.2 to 1331.7-183.46,795
2191322.8Llama 4 Scout 17b 16e InstructMeta1318.1 to 1327.5-184.030,273
2201322.1GPT 4.1 Nano 2025 04 14OpenAI1314.3 to 1329.8-184.76,103
2211321.4Claude 3 Opus 20240229Anthropic1318.4 to 1324.5-185.4194,909
2221320.6Ring Flash 2.0Ant Group1313.4 to 1327.8-186.27,135
2231320.2Step 1o Turbo 202506StepFun1313.4 to 1327.0-186.69,032
2241319.4Glm 4 PlusZ.ai1314.4 to 1324.3-187.426,126
2251318.2Llama 3.3 70b InstructMeta1314.7 to 1321.7-188.654,720
2261318.0Gemma 3n E4b ItGoogle1312.9 to 1323.2-188.822,569
2271318.0Qwen Max 0919Alibaba1312.3 to 1323.7-188.816,478
2281317.7GPT 4o Mini 2024 07 18OpenAI1314.1 to 1321.2-189.168,709
2291317.2GPT Oss 20bOpenAI1310.9 to 1323.6-189.610,620
2301315.6Nvidia Nemotron 3 Nano 30b A3b Bf16NVIDIA1310.0 to 1321.1-191.215,496
2311314.9Qwen2.5 Plus 1127Alibaba1308.5 to 1321.2-191.910,187
2331314.0Mistral Large 2407Mistral AI1310.1 to 1317.9-192.845,459
2341312.7GPT 4 0125 PreviewOpenAI1308.6 to 1316.8-194.193,439
2351312.3GPT 4 1106 PreviewOpenAI1308.4 to 1316.2-194.5100,105
2361311.1Hunyuan Standard 2025 02 10Tencent1301.5 to 1320.8-195.73,904
2371309.2Gemini 1.5 Flash 002Google1305.0 to 1313.4-197.634,902
2381308.4Grok 2 Mini 2024 08 13xAI1304.7 to 1312.1-198.452,567
2391307.4Granite 4.1 8bIBM1297.3 to 1317.5-199.44,064
2401307.1DeepSeek V2.5DeepSeek1302.4 to 1311.8-199.724,572
2421306.1MercuryInception AI1292.3 to 1319.9-200.71,951
2431305.6Olmo 3 32b ThinkAi21297.4 to 1313.7-201.25,938
2441305.3Mistral Large 2411Mistral AI1300.8 to 1309.7-201.528,073
2451304.1Magistral Medium 2506Mistral AI1297.7 to 1310.6-202.711,624
2461303.3Mistral Small 3.1 24b Instruct 2503Mistral AI1298.8 to 1307.8-203.533,194
2471303.3Gemma 3 4b ItGoogle1294.0 to 1312.6-203.54,171
2481302.8Qwen2.5 72b InstructAlibaba1298.7 to 1306.9-204.039,406
2491298.9Llama 3.1 Nemotron 70b InstructNVIDIA1291.2 to 1306.7-207.97,140
2501294.0Hunyuan Large VisionTencent1284.9 to 1303.1-212.85,369
2511293.1Llama 3.1 70b InstructMeta1289.4 to 1296.9-213.755,240
2521290.0Amazon Nova Pro V1.0Amazon1285.4 to 1294.6-216.824,745
2541288.9Gemma 2 27b ItGoogle1285.5 to 1292.3-217.975,754
2561287.0Ibm Granite H SmallIBM1278.6 to 1295.4-219.85,682
2571286.7GPT 4 0314OpenAI1282.0 to 1291.5-220.154,173
2581286.2Gemini 1.5 Flash 001Google1281.7 to 1290.7-220.662,833
2591285.9Llama 3.1 Nemotron 51b InstructNVIDIA1276.0 to 1295.9-220.93,749
2601285.9Llama 3.1 Tulu 3 70bAi21275.5 to 1296.4-220.92,846
2611285.0Olmo 3.1 32b ThinkAi21277.8 to 1292.2-221.88,497
2621280.6Claude 3 Sonnet 20240229Anthropic1276.6 to 1284.6-226.2109,284
2641276.5Nemotron 4 340b InstructNVIDIA1271.2 to 1281.9-230.319,659
2651275.9Llama 3 70b InstructMeta1272.4 to 1279.5-230.9156,876
2661275.8Command R Plus 08 2024Cohere1269.2 to 1282.4-231.09,866
2671275.0GPT 4 0613OpenAI1270.9 to 1279.1-231.888,723
2681274.1Mistral Small 24b Instruct 2501Mistral AI1268.2 to 1280.0-232.714,681
2691273.0Glm 4 0520Z.ai1266.0 to 1280.0-233.89,788
2711270.4Qwen2.5 Coder 32b InstructAlibaba1262.2 to 1278.5-236.45,432
2721266.9C4ai Aya Expanse 32bCohere1262.0 to 1271.8-239.927,124
2731266.4Gemma 2 9b ItGoogle1262.6 to 1270.2-240.454,611
2741264.5DeepSeek Coder V2DeepSeek1258.3 to 1270.8-242.315,147
2751261.1Qwen2 72b InstructAlibaba1256.2 to 1266.1-245.737,325
2761261.1Command R PlusCohere1256.8 to 1265.4-245.777,554
2771260.9Claude 3 Haiku 20240307Anthropic1257.1 to 1264.6-245.9117,701
2781260.3Amazon Nova Lite V1.0Amazon1255.2 to 1265.5-246.519,372
2791258.6Gemini 1.5 Flash 8b 001Google1254.3 to 1262.9-248.235,558
2801256.0Phi 4Microsoft1251.4 to 1260.6-250.824,126
2811251.3Olmo 2 0325 32b InstructAi21240.5 to 1262.1-255.53,334
2821249.7Command R 08 2024Cohere1243.1 to 1256.3-257.110,140
2831241.7Mistral Large 2402Mistral AI1237.0 to 1246.4-265.162,436
2841240.6Amazon Nova Micro V1.0Amazon1235.4 to 1245.7-266.219,364
2861237.3Ministral 8b 2410Mistral AI1228.2 to 1246.4-269.54,781
2871235.7Gemini Pro Dev ApiGoogle1228.4 to 1243.1-271.118,354
2881233.5Qwen1.5 110b ChatAlibaba1227.9 to 1239.0-273.326,195
2891233.2Hunyuan Standard 256kTencent1221.5 to 1245.0-273.62,728
2911232.7Qwen1.5 72b ChatAlibaba1227.4 to 1238.0-274.139,302
2921228.7Mixtral 8x22b Instruct V0.1Mistral AI1224.2 to 1233.3-278.151,416
2931226.1Command RCohere1221.3 to 1230.9-280.754,036
2951224.2GPT 3.5 Turbo 0125OpenAI1219.5 to 1228.9-282.666,207
2961223.0Llama 3 8b InstructMeta1219.2 to 1226.7-283.8104,642
2971222.8C4ai Aya Expanse 8bCohere1215.8 to 1229.8-284.09,818
2981222.4Gemini ProGoogle1210.7 to 1234.2-284.46,390
2991222.0Mistral MediumMistral AI1216.5 to 1227.5-284.834,550
3001220.4Llama 3.1 Tulu 3 8bAi21209.7 to 1231.1-286.42,896
3031211.3Llama 3.1 8b InstructMeta1207.1 to 1215.4-295.549,605
3041207.8Granite 3.1 8b InstructIBM1196.7 to 1218.9-299.03,090
3051203.2Qwen1.5 32b ChatAlibaba1197.0 to 1209.3-303.621,741
3061202.7GPT 3.5 Turbo 1106OpenAI1194.0 to 1211.5-304.116,619
3071199.8Gemma 2 2b ItGoogle1195.7 to 1203.9-307.046,616
3081197.2Phi 3 Medium 4k InstructMicrosoft1192.0 to 1202.4-309.625,055
3091196.3Mixtral 8x7b Instruct V0.1Mistral AI1192.1 to 1200.6-310.573,503
3121190.3Qwen1.5 14b ChatAlibaba1183.2 to 1197.4-316.517,839
3131183.9Wizardlm 70bMicrosoft1174.5 to 1193.4-322.98,214
3141183.9DeepSeek LLM 67b ChatDeepSeek1172.4 to 1195.4-322.94,932
3161182.0Granite 3.0 8b InstructIBM1173.4 to 1190.7-324.86,638
3191181.5Gemma 1.1 7b ItGoogle1175.5 to 1187.6-325.323,893
3211178.3Granite 3.1 2b InstructIBM1167.1 to 1189.5-328.53,188
3261170.4Phi 3 Small 8k InstructMicrosoft1164.4 to 1176.3-336.417,766
3271170.0Llama 2 70b ChatMeta1164.5 to 1175.5-336.838,492
3291166.2Llama 3.2 3b InstructMeta1158.5 to 1173.9-340.67,936
3311155.7Granite 3.0 2b InstructIBM1147.3 to 1164.1-351.16,837
3321154.6Qwq 32b PreviewAlibaba1143.2 to 1166.1-352.23,231
3331154.0Llama2 70b Steerlm ChatNVIDIA1141.4 to 1166.6-352.83,585
3371148.7Mistral 7b Instruct V0.2Mistral AI1142.0 to 1155.3-358.119,402
3381148.6Wizardlm 13bMicrosoft1139.4 to 1157.8-358.27,044
3401143.1Qwen1.5 7b ChatAlibaba1133.3 to 1153.0-363.74,737
3411142.4Phi 3 Mini 4k Instruct June 2024Microsoft1135.9 to 1148.8-364.412,297
3421140.8Llama 2 13b ChatMeta1134.1 to 1147.5-366.019,174
3441138.2Qwen 14b ChatAlibaba1127.3 to 1149.1-368.64,964
3451137.7Palm 2Google1128.3 to 1147.0-369.18,554
3461136.8Gemma 7b ItGoogle1127.3 to 1146.3-370.08,925
3471136.0Codellama 34b InstructMeta1127.1 to 1144.9-370.87,366
3491128.8Phi 3 Mini 128k InstructMicrosoft1121.4 to 1136.3-378.020,685
3501127.5Phi 3 Mini 4k InstructMicrosoft1121.1 to 1133.9-379.320,118
3541118.5Codellama 70b InstructMeta1100.3 to 1136.6-388.31,143
3551115.5Gemma 1.1 2b ItGoogle1107.8 to 1123.3-391.310,854
3581110.5Llama 3.2 1b InstructMeta1102.7 to 1118.4-396.38,045
3591109.2Mistral 7b InstructMistral AI1100.0 to 1118.5-397.68,977
3601107.3Llama 2 7b ChatMeta1100.3 to 1114.4-399.514,148
3611092.6Gemma 2b ItGoogle1081.1 to 1104.1-414.24,780
3621089.9Qwen1.5 4b ChatAlibaba1080.6 to 1099.2-416.97,597
3631073.2Olmo 7b InstructAi21062.0 to 1084.4-433.66,328
375973.3Llama 13bMeta957.5 to 989.1-533.52,391

Provider frontier

The leading model per organization.

LMArena rating, style-controlled text category, published July 16, 2026. Gap to leader is calculated inside this one snapshot.
RankLMArena ratingOrganizationLeading modelReported intervalGap to leaderBattles
011507.5AnthropicClaude Fable 51500.2 to 1514.7Leader8,817
021493.1MetaMuse Spark 1.11484.9 to 1501.3-14.45,732
031486.3OpenAIGPT-5.6 Sol X-High1476.8 to 1495.8-21.24,113
041485.8GoogleGemini 3 Pro1481.9 to 1489.6-21.741,283
051474.2xAIGrok 4.20 Beta 11469.5 to 1478.8-33.326,844
061457.5DeepSeekDeepSeek V4 Pro1453.1 to 1462.0-50.043,062
071427.3MistralMistral Medium 3.51420.7 to 1433.9-80.211,017

Across the boards

Model Gauntlet mirrors named authorities as separate labeled views, each with its own subject, source, date, and license. These columns are not combined and no model is ranked across them.

LMArena

Human preference

  1. 01Claude Fable 51507.5
  2. 02Muse Spark 1.11493.1
  3. 03GPT-5.6 Sol X-High1486.3
  4. 04Gemini 3 Pro1485.8
  5. 05Grok 4.20 Beta 11474.2
  6. 06DeepSeek V4 Pro1457.5
  7. 07Mistral Medium 3.51427.3

LMArena leaderboard dataset, CC-BY-4.0. Published July 16, 2026.

Aider polyglot

Code editing systems

Aider ranks full model-and-editor-system configurations on polyglot code editing, not bare models. These pass rates are not comparable with LMArena preference scores.

  1. 01gpt-5 (high)88.0%
  2. 02gpt-5 (medium)86.7%
  3. 03o3-pro (high)84.9%
  4. 04gemini-2.5-pro-preview-06-05 (32k think)83.1%
  5. 05o3 (high)81.3%
  6. 06gpt-5 (low)81.3%
  7. 07grok-4 (high)79.6%
  8. 08gemini-2.5-pro-preview-06-05 (default think)79.1%
  9. 09o3 (high) + gpt-4.178.2%
  10. 10Gemini 2.5 Pro Preview 05-0676.9%

Aider polyglot leaderboard, Aider-AI/aider, Apache-2.0. Source last updated May 22, 2026.

EQ-Bench 3

Emotional intelligence

  1. 01claude-opus-4-71461
  2. 02claude-sonnet-4-61384
  3. 03claude-opus-4-61383
  4. 04HiveLabsAI/hivemind-32b-preview1351
  5. 05deepseek/deepseek-v4-pro1330
  6. 06openai/gpt-5.41328
  7. 07openai/gpt-5.51328
  8. 08moonshotai/kimi-k2.61319
  9. 09gpt-5.1-2025-11-131314
  10. 10gpt-5.21314

EQ-Bench 3 canonical Elo leaderboard, EQ-bench/eqbench3, MIT. Source last updated May 10, 2026.

Category boards

This page is the overall text category. LMArena also publishes per-category arenas, each mirrored on its own board with a frontier race by organization.

Leaderboard questions

Which organization leads this leaderboard right now?

According to the LMArena snapshot published July 16, 2026, Anthropic leads with Claude Fable 5 at 1507.5, 14.4 points ahead of Meta in the style-controlled text category.

Is this a combined ranking across many benchmarks?

No. Both tables report one LMArena rating published July 16, 2026: the model rankings list every resolved model with LMArena's official rank, and the provider table shows the highest-scoring model per organization. Model Gauntlet publishes no composite or consensus score.

Why do some rank numbers skip?

The rank column is LMArena's official rank from the published snapshot. Model Gauntlet lists only rows that resolve to a reviewed model identity, so a skipped number is a source row without a reviewed identity match. 324 of the 375 officially ranked rows resolve in this snapshot.

Why does each score show an interval?

LMArena publishes a reported 95% confidence interval with each rating. Where two intervals overlap, this one snapshot does not separate those ranks decisively, so treat small gaps with caution.

How current are these numbers?

The pinned source snapshot was published July 16, 2026 and stays labeled with that date until a newer snapshot passes human review. Nothing on this page is described as live.

Which boards does Model Gauntlet track?

Model Gauntlet mirrors three named authorities as separate labeled views. The LMArena leaderboard dataset covers human preference and fills the tables above, published July 16, 2026. The Aider polyglot leaderboard covers code editing systems, and EQ-Bench 3 covers emotional intelligence; both appear in the Across the boards section below. Each carries its own source, date, and license, and Model Gauntlet publishes no score that combines them.