LMArena Coding Arena Leaderboard
The latest AI coding model leaderboard based on LMArena Coding Arena anonymous user voting. Covers Elo scores, confidence intervals, and vote counts for Claude, GPT, Gemini, DeepSeek, Qwen, and more.
Top Model
Claude Fable 5
Top Score
1553.00
Model Count
380
Data version
2026年08月06日
Data source: LM Arena
About This Leaderboard
This leaderboard ranks AI models by coding ability. Data comes from LMArena (formerly LMSYS Chatbot Arena)'s Coding sub-track, evaluated through anonymous blind testing by real users on programming tasks.
Methodology Overview
Blind testing: Users submit coding questions, two anonymous models generate code answers, and users vote for the better response — eliminating brand bias.
Elo scoring: Uses the Bradley-Terry model to calculate Elo scores. Higher scores mean users more frequently prefer that model's code solutions.
Broad scenario coverage: Testing spans code generation, bug fixing, algorithm implementation, code explanation, and more real-world programming scenarios.
DataLearner provides in-depth analysis on top of the raw data, linking leaderboard models to the DataLearner model database so you can quickly access model details, API pricing, benchmark scores, and more.
Ranking Table
| Rank | Model | Score | 95% CI | Votes | Organization | License |
|---|---|---|---|---|---|---|
Claude Fable 5Anthropic | 1553.00 | +/-9 | 5,234 | Anthropic | Proprietary | |
Opus 4.7 (thinking)Anthropic | 1552.00 | +/-6 | 16,303 | Anthropic | Proprietary | |
Claude Opus 4.6 (thinking)Anthropic | 1551.00 | +/-6 | 18,015 | Anthropic | Proprietary | |
| 4 | Claude Opus 4.6Anthropic | 1547.00 | +/-6 | 20,444 | Anthropic | Proprietary |
| 5 | Opus 4.7Anthropic | 1546.00 | +/-6 | 16,534 | Anthropic | Proprietary |
| 6 | kimi-k3-maxMoonshot | 1542.00 | +/-12 | 2,611 | Moonshot | Kimi K3 license |
| 7 | Claude Opus 4.8 (thinking)Anthropic | 1535.00 | +/-7 | 10,509 | Anthropic | Proprietary |
| 8 | qwen3.8-maxAlibaba | 1532.00 | +/-16 | 1,411 | Alibaba | Proprietary |
| 9 | Muse Spark 1.1Facebook AI研究实验室 | 1532.00 | +/-10 | 4,199 | Facebook AI研究实验室 | Proprietary |
| 10 | Claude Opus 4 (thinking-32k)Anthropic | 1530.00 | +/-7 | 7,603 | Anthropic | Proprietary |
| 11 | claude-opus-5-highAnthropic | 1530.00 | +/-10 | 4,043 | Anthropic | Proprietary |
| 12 | Claude Sonnet 4.6Anthropic | 1528.00 | +/-6 | 17,610 | Anthropic | Proprietary |
| 13 | claude-opus-5-maxAnthropic | 1527.00 | +/-14 | 1,910 | Anthropic | Proprietary |
| 14 | Claude Opus 4.8Anthropic | 1527.00 | +/-7 | 10,721 | Anthropic | Proprietary |
| 15 | muse-spark-1.2 (xHigh)Meta | 1526.00 | +/-25 | 577 | Meta | Proprietary |
| 16 | Muse SparkFacebook AI研究实验室 | 1526.00 | +/-10 | 3,744 | Facebook AI研究实验室 | Proprietary |
| 17 | Qwen3.7-Max-Preview阿里巴巴 | 1526.00 | +/-18 | 1,112 | 阿里巴巴 | Proprietary |
| 18 | gpt-5.6-sol-xhighOpenAI | 1525.00 | +/-10 | 3,665 | OpenAI | Proprietary |
| 19 | Claude Sonnet 4.5 (high)Anthropic | 1525.00 | +/-9 | 5,698 | Anthropic | Proprietary |
| 20 | gemini-3.6-flashGoogle | 1524.00 | +/-11 | 3,364 | Proprietary | |
| 21 | Claude Opus 4Anthropic | 1523.00 | +/-6 | 17,181 | Anthropic | Proprietary |
| 22 | GPT-5.4 (high)OpenAI | 1521.00 | +/-6 | 16,326 | OpenAI | Proprietary |
| 23 | Gemini 3.1 Pro PreviewGoogle Deep Mind | 1521.00 | +/-5 | 25,189 | Google Deep Mind | Proprietary |
| 24 | gpt-5.6-terra-xhighOpenAI | 1521.00 | +/-10 | 3,745 | OpenAI | Proprietary |
| 25 | 1520.00 | +/-10 | 4,345 | xAI | Proprietary | |
| 26 | Claude Sonnet 4.5 (thinking-32k)Anthropic | 1520.00 | +/-5 | 19,186 | Anthropic | Proprietary |
| 27 | GPT-5.5 (high)OpenAI | 1519.00 | +/-6 | 14,532 | OpenAI | Proprietary |
| 28 | mimo-v2.5-proXiaomi | 1519.00 | +/-7 | 13,199 | Xiaomi | MIT |
| 29 | Gemini 3 ProGoogle Deep Mind | 1519.00 | +/-7 | 8,561 | Google Deep Mind | Proprietary |
| 30 | GLM 5.1智谱AI | 1517.00 | +/-7 | 10,330 | 智谱AI | MIT |
| 31 | GPT-5.2 Chat (0210)OpenAI | 1515.00 | +/-7 | 9,036 | OpenAI | Proprietary |
| 32 | Kimi K2.6Moonshot AI | 1515.00 | +/-7 | 10,310 | Moonshot AI | Modified MIT |
| 33 | GPT-5.5 InstantOpenAI | 1515.00 | +/-8 | 7,461 | OpenAI | Proprietary |
| 34 | GPT-5.4OpenAI | 1514.00 | +/-6 | 17,549 | OpenAI | Proprietary |
| 35 | ERNIE-5.1-Preview百度 | 1514.00 | +/-7 | 10,223 | 百度 | Proprietary |
| 36 | Claude Sonnet 4.5Anthropic | 1513.00 | +/-5 | 19,024 | Anthropic | Proprietary |
| 37 | DOLA Seed 2.0 Pro字节跳动Seed团队 | 1513.00 | +/-6 | 20,619 | 字节跳动Seed团队 | Proprietary |
| 38 | Qwen3.5 Max Preview阿里巴巴 | 1513.00 | +/-8 | 6,000 | 阿里巴巴 | Proprietary |
| 39 | Opus 4.1 (thinking-16k)Anthropic | 1513.00 | +/-6 | 9,825 | Anthropic | Proprietary |
| 40 | 1511.00 | +/-6 | 17,074 | xAI | Proprietary | |
| 41 | gemini-3.5-flash-highGoogle | 1510.00 | +/-8 | 6,718 | Proprietary | |
| 42 | GPT-5.5OpenAI | 1510.00 | +/-6 | 15,231 | OpenAI | Proprietary |
| 43 | 1509.00 | +/-6 | 16,757 | xAI | Proprietary | |
| 44 | Qwen3.6-Max-Preview阿里巴巴 | 1508.00 | +/-16 | 1,533 | 阿里巴巴 | Proprietary |
| 45 | Gemini 3.0 FlashGoogle Deep Mind | 1508.00 | +/-8 | 6,374 | Google Deep Mind | Proprietary |
| 46 | 1508.00 | +/-8 | 7,136 | xAI | Proprietary | |
| 47 | gemini-3.5-flash-liteGoogle | 1507.00 | +/-11 | 3,137 | Proprietary | |
| 48 | gemini-3.5-flash-mediumGoogle | 1507.00 | +/-8 | 6,365 | Proprietary | |
| 49 | glm-5.2-maxZ.ai | 1507.00 | +/-8 | 6,940 | Z.ai | MIT |
| 50 | Qwen3.7-Plus阿里巴巴 | 1505.00 | +/-8 | 8,269 | 阿里巴巴 | Proprietary |
| 51 | Opus 4.1Anthropic | 1505.00 | +/-5 | 15,507 | Anthropic | Proprietary |
| 52 | Kimi K2.5 InstantMoonshot AI | 1504.00 | +/-14 | 1,797 | Moonshot AI | Modified MIT |
| 53 | mimo-v2-proXiaomi | 1503.00 | +/-8 | 6,872 | Xiaomi | Proprietary |
| 54 | longcat-flash-chat-2602-expMeituan | 1502.00 | +/-8 | 7,668 | Meituan | Proprietary |
| 55 | Kimi K2 ThinkingMoonshot AI | 1502.00 | +/-5 | 18,836 | Moonshot AI | Modified MIT |
| 56 | DeepSeek-V4-ProDeepSeek-AI | 1501.00 | +/-6 | 15,292 | DeepSeek-AI | MIT |
| 57 | Hy3腾讯AI实验室 | 1500.00 | +/-18 | 1,133 | 腾讯AI实验室 | Apache 2.0 |
| 58 | 1499.00 | +/-7 | 10,050 | MiniMaxAI | MiniMax Community License | |
| 59 | Claude Opus 4 (thinking-16k)Anthropic | 1499.00 | +/-8 | 6,665 | Anthropic | Proprietary |
| 60 | 1499.00 | +/-6 | 14,988 | xAI | Proprietary | |
| 61 | Gemma 4 31BDeepMind | 1498.00 | +/-15 | 1,354 | DeepMind | Apache 2.0 |
| 62 | gpt-5.6-luna-xhighOpenAI | 1498.00 | +/-10 | 3,936 | OpenAI | Proprietary |
| 63 | GLM-5智谱AI | 1497.00 | +/-7 | 7,166 | 智谱AI | MIT |
| 64 | GPT-5.3 ChatOpenAI | 1497.00 | +/-7 | 8,671 | OpenAI | Proprietary |
| 65 | GPT-5.4 mini (high)OpenAI | 1497.00 | +/-6 | 16,094 | OpenAI | Proprietary |
| 66 | Qwen 3.6 Plus Preview阿里巴巴 | 1495.00 | +/-6 | 13,660 | 阿里巴巴 | Proprietary |
| 67 | InklingThinking Machines Lab | 1493.00 | +/-10 | 3,573 | Thinking Machines Lab | Apache 2.0 |
| 68 | 1492.00 | +/-6 | 15,554 | xAI | Proprietary | |
| 69 | Qwen3.5-397B-A17B阿里巴巴 | 1492.00 | +/-6 | 18,329 | 阿里巴巴 | Apache 2.0 |
| 70 | Gemini 3.0 Flash (minimal)Google Deep Mind | 1491.00 | +/-5 | 22,449 | Google Deep Mind | Proprietary |
| 71 | mimo-v2.5Xiaomi | 1491.00 | +/-7 | 12,790 | Xiaomi | MIT |
| 72 | GPT-5.1 Pro (high)OpenAI | 1491.00 | +/-7 | 8,185 | OpenAI | Proprietary |
| 73 | ERNIE 5.0 (0110)百度 | 1490.00 | +/-7 | 8,547 | 百度 | Proprietary |
| 74 | GPT-5.2 Pro (high)OpenAI | 1489.00 | +/-6 | 11,656 | OpenAI | Proprietary |
| 75 | deepseek-v4-pro-high-previewDeepSeek | 1489.00 | +/-6 | 14,268 | DeepSeek | MIT |
| 76 | GLM-5V-Turbo智谱AI | 1488.00 | +/-12 | 2,576 | 智谱AI | Proprietary |
| 77 | 1488.00 | +/-6 | 15,091 | xAI | Proprietary | |
| 78 | mimo-v2-omniXiaomi | 1487.00 | +/-9 | 5,477 | Xiaomi | Proprietary |
| 79 | Kimi K2 Thinking (thinking-turbo)Moonshot AI | 1486.00 | +/-6 | 14,758 | Moonshot AI | Modified MIT |
| 80 | amazon-nova-experimental-chat-26-02-10Amazon | 1486.00 | +/-20 | 841 | Amazon | Proprietary |
| 81 | GLM-4.7智谱AI | 1485.00 | +/-12 | 2,410 | 智谱AI | MIT |
| 82 | DeepSeek-V4-FlashDeepSeek-AI | 1483.00 | +/-6 | 14,249 | DeepSeek-AI | MIT |
| 83 | GPT-5.2OpenAI | 1482.00 | +/-5 | 20,657 | OpenAI | Proprietary |
| 84 | Qwen3 Max (Preview)阿里巴巴 | 1482.00 | +/-8 | 5,368 | 阿里巴巴 | Proprietary |
| 85 | Gemma 4 26B A4BDeepMind | 1481.00 | +/-15 | 1,357 | DeepMind | Apache 2.0 |
| 86 | deepseek-v4-flash-high-previewDeepSeek | 1479.00 | +/-6 | 14,203 | DeepSeek | MIT |
| 87 | amazon-nova-experimental-chat-26-01-10Amazon | 1479.00 | +/-21 | 730 | Amazon | Proprietary |
| 88 | Haiku 4.5Anthropic | 1479.00 | +/-5 | 28,992 | Anthropic | Proprietary |
| 89 | Mistral Medium 3.5MistralAI | 1479.00 | +/-11 | 3,046 | MistralAI | Modified MIT |
| 90 | 1479.00 | +/-6 | 16,121 | MiniMaxAI | Modified MIT | |
| 91 | DeepSeek V3.2-Exp (thinking)DeepSeek-AI | 1475.00 | +/-7 | 8,493 | DeepSeek-AI | MIT |
| 92 | DeepSeek V3.2-Exp (thinking)DeepSeek-AI | 1475.00 | +/-13 | 1,914 | DeepSeek-AI | MIT |
| 93 | Nemotron 3 UltraNVIDIA | 1474.00 | +/-12 | 2,910 | NVIDIA | OpenMDW-1.1 |
| 94 | GPT-5.1OpenAI | 1474.00 | +/-7 | 9,087 | OpenAI | Proprietary |
| 95 | qwen3-max-2025-09-23Alibaba | 1474.00 | +/-13 | 2,039 | Alibaba | Proprietary |
| 96 | longcat-flash-chatMeituan | 1474.00 | +/-13 | 2,233 | Meituan | MIT |
| 97 | Claude Sonnet 4 (thinking-32k)Anthropic | 1473.00 | +/-8 | 6,401 | Anthropic | Proprietary |
| 98 | Qwen3-235B-A22B-2507阿里巴巴 | 1472.00 | +/-5 | 21,199 | 阿里巴巴 | Apache 2.0 |
| 99 | ERNIE 5.0 Preview (1203)百度 | 1472.00 | +/-13 | 1,952 | 百度 | Proprietary |
| 100 | DeepSeek V3.2DeepSeek-AI | 1470.00 | +/-6 | 10,556 | DeepSeek-AI | MIT |
| 101 | GPT-4o(2025-03-27)OpenAI | 1469.00 | +/-5 | 15,841 | OpenAI | Proprietary |
| 102 | Mistral Large 3MistralAI | 1469.00 | +/-6 | 13,780 | MistralAI | Apache 2.0 |
| 103 | GPT-5-Pro (high)OpenAI | 1468.00 | +/-8 | 6,356 | OpenAI | Proprietary |
| 104 | Kimi K2 0905Moonshot AI | 1468.00 | +/-13 | 2,240 | Moonshot AI | Modified MIT |
| 105 | Gemini 2.5 ProGoogle Deep Mind | 1465.00 | +/-4 | 26,439 | Google Deep Mind | Proprietary |
| 106 | DeepSeek V3.2-ExpDeepSeek-AI | 1465.00 | +/-12 | 2,491 | DeepSeek-AI | MIT |
| 107 | Qwen3-VL-235B-A22B-Instruct阿里巴巴 | 1465.00 | +/-13 | 2,314 | 阿里巴巴 | Apache 2.0 |
| 108 | DeepSeek-R1-0528DeepSeek-AI | 1464.00 | +/-11 | 2,725 | DeepSeek-AI | MIT |
| 109 | Claude Opus 4Anthropic | 1464.00 | +/-7 | 7,892 | Anthropic | Proprietary |
| 110 | GPT-5.2 Chat (0210)OpenAI | 1464.00 | +/-8 | 5,977 | OpenAI | Proprietary |
| 111 | DeepSeek-V3.1 Terminus (thinking)DeepSeek-AI | 1463.00 | +/-24 | 635 | DeepSeek-AI | MIT |
| 112 | 1461.00 | +/-6 | 13,507 | xAI | Proprietary | |
| 113 | GPT-5.4 nano (high)OpenAI | 1461.00 | +/-6 | 16,436 | OpenAI | Proprietary |
| 114 | Kimi K2Moonshot AI | 1460.00 | +/-8 | 5,235 | Moonshot AI | Modified MIT |
| 115 | hunyuan-hy3-previewTencent | 1460.00 | +/-14 | 1,942 | Tencent | tencent-hunyuan-community |
| 116 | OpenAI o3OpenAI | 1459.00 | +/-6 | 11,733 | OpenAI | Proprietary |
| 117 | Qwen3.5-122B-A10B阿里巴巴 | 1459.00 | +/-7 | 7,740 | 阿里巴巴 | Apache 2.0 |
| 118 | 1459.00 | +/-16 | 1,247 | xAI | Proprietary | |
| 119 | GPT-4.5OpenAI | 1459.00 | +/-13 | 1,939 | OpenAI | Proprietary |
| 120 | GLM-4.6智谱AI | 1459.00 | +/-7 | 7,469 | 智谱AI | MIT |
| 121 | Gemini 3.1 Flash-LiteGoogle Deep Mind | 1457.00 | +/-6 | 16,599 | Google Deep Mind | Proprietary |
| 122 | DeepSeek-V3.1 (thinking)DeepSeek-AI | 1457.00 | +/-13 | 1,902 | DeepSeek-AI | MIT |
| 123 | Qwen3-Coder-480B-A35B阿里巴巴 | 1457.00 | +/-9 | 4,844 | 阿里巴巴 | Apache 2.0 |
| 124 | GPT-4.1OpenAI | 1456.00 | +/-7 | 9,301 | OpenAI | Proprietary |
| 125 | Magistral-Medium-2506MistralAI | 1455.00 | +/-5 | 20,986 | MistralAI | Proprietary |
| 126 | Qwen3-VL-235B-A22B-Instruct (thinking)阿里巴巴 | 1455.00 | +/-14 | 1,627 | 阿里巴巴 | Apache 2.0 |
| 127 | GLM-4.5智谱AI | 1454.00 | +/-9 | 4,765 | 智谱AI | MIT |
| 128 | Claude3-Sonnet (thinking-32k)Anthropic | 1452.00 | +/-8 | 6,189 | Anthropic | Proprietary |
| 129 | Step 3.5 FlashStepFunAI | 1451.00 | +/-6 | 14,275 | StepFunAI | Apache 2.0 |
| 130 | Qwen3.5-27B阿里巴巴 | 1450.00 | +/-7 | 7,446 | 阿里巴巴 | Apache 2.0 |
| 131 | Claude Sonnet 4Anthropic | 1449.00 | +/-7 | 7,386 | Anthropic | Proprietary |
| 132 | DeepSeek-V3.1DeepSeek-AI | 1448.00 | +/-12 | 2,624 | DeepSeek-AI | MIT |
| 133 | mimo-v2-flash (non-thinking)Xiaomi | 1446.00 | +/-6 | 11,911 | Xiaomi | MIT |
| 134 | qwen3-235b-a22b-no-thinkingAlibaba | 1446.00 | +/-8 | 6,967 | Alibaba | Apache 2.0 |
| 135 | Qwen3-Next阿里巴巴 | 1445.00 | +/-9 | 4,790 | 阿里巴巴 | Apache 2.0 |
| 136 | DeepSeek-R1DeepSeek-AI | 1445.00 | +/-12 | 2,317 | DeepSeek-AI | MIT |
| 137 | 1444.00 | +/-7 | 10,772 | MiniMaxAI | Modified MIT | |
| 138 | 1443.00 | +/-8 | 5,394 | xAI | Proprietary | |
| 139 | trinity-large-preview Apache 2.0 | 1443.00 | +/-8 | 7,530 | — | — |
| 140 | qwen3-235b-a22b-thinking-2507Alibaba | 1442.00 | +/-15 | 1,612 | Alibaba | Apache 2.0 |
| 141 | 1439.00 | +/-10 | 3,409 | MiniMaxAI | MIT | |
| 142 | Qwen3-30B-A3B-2507阿里巴巴 | 1439.00 | +/-9 | 4,656 | 阿里巴巴 | Apache 2.0 |
| 143 | DeepSeek-V3.1 TerminusDeepSeek-AI | 1439.00 | +/-21 | 778 | DeepSeek-AI | MIT |
| 144 | 1437.00 | +/-9 | 3,945 | xAI | Proprietary | |
| 145 | hunyuan-vision-1.5-thinkingTencent | 1437.00 | +/-27 | 436 | Tencent | Proprietary |
| 146 | Step 3.5 FlashStepFunAI | 1437.00 | +/-6 | 16,519 | StepFunAI | Proprietary |
| 147 | 1435.00 | +/-7 | 8,144 | xAI | Proprietary | |
| 148 | Claude 3.5 SonnetAnthropic | 1435.00 | +/-6 | 14,955 | Anthropic | Proprietary |
| 149 | OpenAI o3-mini (high)OpenAI | 1435.00 | +/-12 | 2,596 | OpenAI | Proprietary |
| 150 | Qwen3.5-35B-A3B阿里巴巴 | 1435.00 | +/-7 | 7,896 | 阿里巴巴 | Apache 2.0 |
| 151 | GPT-4.1 miniOpenAI | 1434.00 | +/-7 | 6,911 | OpenAI | Proprietary |
| 152 | amazon-nova-experimental-chat-12-10Amazon | 1433.00 | +/-21 | 704 | Amazon | Proprietary |
| 153 | mistral-medium-2505Mistral | 1433.00 | +/-8 | 5,891 | Mistral | Proprietary |
| 154 | OpenAI o1OpenAI | 1433.00 | +/-10 | 3,973 | OpenAI | Proprietary |
| 155 | Qwen3-235B-A22B阿里巴巴 | 1433.00 | +/-9 | 4,340 | 阿里巴巴 | Apache 2.0 |
| 156 | OpenAI o4 - miniOpenAI | 1433.00 | +/-7 | 8,705 | OpenAI | Proprietary |
| 157 | ERNIE 5.0百度 | 1432.00 | +/-19 | 915 | 百度 | Proprietary |
| 158 | GPT-5-mini (high)OpenAI | 1431.00 | +/-8 | 5,497 | OpenAI | Proprietary |
| 159 | mimo-v2-flash (thinking)Xiaomi | 1431.00 | +/-12 | 2,437 | Xiaomi | MIT |
| 160 | Claude3-SonnetAnthropic | 1430.00 | +/-7 | 7,132 | Anthropic | Proprietary |
| 161 | DeepSeek-V3-0324DeepSeek-AI | 1429.00 | +/-7 | 8,358 | DeepSeek-AI | MIT |
| 162 | Gemini 2.5 Flash-Preview-09-2025Google Deep Mind | 1428.00 | +/-7 | 6,835 | Google Deep Mind | Proprietary |
| 163 | GLM-4.5-Air智谱AI | 1426.00 | +/-8 | 6,099 | 智谱AI | MIT |
| 164 | GLM-4.7-Flash智谱AI | 1424.00 | +/-11 | 2,673 | 智谱AI | MIT |
| 165 | Gemini 2.5 FlashGoogle Deep Mind | 1424.00 | +/-4 | 25,761 | Google Deep Mind | Proprietary |
| 166 | Qwen3-Next (thinking)阿里巴巴 | 1421.00 | +/-11 | 2,672 | 阿里巴巴 | Apache 2.0 |
| 167 | GLM-4.6V智谱AI | 1419.00 | +/-25 | 533 | 智谱AI | MIT |
| 168 | amazon-nova-experimental-chat-11-10Amazon | 1419.00 | +/-8 | 5,285 | Amazon | Proprietary |
| 169 | minimax-m1MiniMax | 1416.00 | +/-8 | 6,471 | MiniMax | Apache 2.0 |
| 170 | OpenAI o1OpenAI | 1416.00 | +/-9 | 5,123 | OpenAI | Proprietary |
| 171 | OpenAI o3-miniOpenAI | 1416.00 | +/-6 | 9,450 | OpenAI | Proprietary |
| 172 | trinity-large-thinking Apache 2.0 | 1414.00 | +/-8 | 8,066 | — | — |
| 173 | Mistral-Small-3.2MistralAI | 1412.00 | +/-10 | 3,357 | MistralAI | Apache 2.0 |
| 174 | ling-flash-2.0 AntGroup | 1412.00 | +/-15 | 1,528 | Group | MIT |
| 175 | amazon-nova-experimental-chat-10-20Amazon | 1410.00 | +/-12 | 2,288 | Amazon | Proprietary |
| 176 | nvidia-nemotron-3-super-120b-a12bNvidia | 1410.00 | +/-14 | 1,777 | Nvidia | NVIDIA Open Model |
| 177 | intellect-3 MIT | 1409.00 | +/-19 | 971 | — | — |
| 178 | Step3StepFunAI | 1408.00 | +/-17 | 1,231 | StepFunAI | Apache 2.0 |
| 179 | Qwen3-32B阿里巴巴 | 1407.00 | +/-24 | 513 | 阿里巴巴 | Apache 2.0 |
| 180 | nvidia-llama-3.3-nemotron-super-49b-v1.5Nvidia | 1405.00 | +/-22 | 659 | Nvidia | Nvidia Open |
| 181 | GLM-4.5V智谱AI | 1404.00 | +/-18 | 991 | 智谱AI | MIT |
| 182 | Qwen2.5-Max阿里巴巴 | 1403.00 | +/-8 | 5,094 | 阿里巴巴 | Proprietary |
| 183 | hunyuan-turbos-20250226Tencent | 1400.00 | +/-31 | 275 | Tencent | Proprietary |
| 184 | Hunyuan-T1腾讯AI实验室 | 1399.00 | +/-20 | 802 | 腾讯AI实验室 | Proprietary |
| 185 | Claude 3.5 SonnetAnthropic | 1399.00 | +/-7 | 13,607 | Anthropic | Proprietary |
| 186 | Gemini 2.5 Flash-Lite-Preview-09-2025 (no-thinking)Google Deep Mind | 1397.00 | +/-6 | 9,668 | Google Deep Mind | Proprietary |
| 187 | mercury-2 InceptionAI | 1396.00 | +/-21 | 764 | AI | Proprietary |
| 188 | Nova 2 Lite亚马逊 | 1396.00 | +/-12 | 2,507 | 亚马逊 | Proprietary |
| 189 | hunyuan-turbos-20250416Tencent | 1394.00 | +/-14 | 1,776 | Tencent | Proprietary |
| 190 | llama-3.1-nemotron-ultra-253b-v1Nvidia | 1391.00 | +/-30 | 367 | Nvidia | Nvidia Open Model |
| 191 | ring-flash-2.0 AntGroup | 1391.00 | +/-15 | 1,537 | Group | MIT |
| 192 | GPT OSS 120BOpenAI | 1390.00 | +/-8 | 6,487 | OpenAI | Apache 2.0 |
| 193 | OpenAI o3-mini (high)OpenAI | 1390.00 | +/-10 | 3,292 | OpenAI | Proprietary |
| 194 | amazon-nova-experimental-chat-10-09Amazon | 1390.00 | +/-24 | 550 | Amazon | Proprietary |
| 195 | C4AI Command A (202503)CohereAI | 1389.00 | +/-6 | 10,208 | CohereAI | CC-BY-NC-4.0 |
| 196 | DeepSeek-V3DeepSeek-AI | 1388.00 | +/-10 | 3,280 | DeepSeek-AI | DeepSeek |
| 197 | Magistral-Medium-2506MistralAI | 1387.00 | +/-12 | 2,244 | MistralAI | Proprietary |
| 198 | OpenAI o1-miniOpenAI | 1387.00 | +/-7 | 8,478 | OpenAI | Proprietary |
| 199 | Qwen3-30B-A3B阿里巴巴 | 1387.00 | +/-9 | 4,526 | 阿里巴巴 | Apache 2.0 |
| 200 | 1386.00 | +/-9 | 4,251 | xAI | Proprietary | |
| 201 | Claude 3.5 HaikuAnthropic | 1385.00 | +/-6 | 11,234 | Anthropic | Proprietary |
| 202 | 1385.00 | +/-15 | 1,545 | MiniMaxAI | Apache 2.0 | |
| 203 | QwQ-32B阿里巴巴 | 1384.00 | +/-9 | 4,040 | 阿里巴巴 | Apache 2.0 |
| 204 | olmo-3.1-32b-instructAi2 | 1384.00 | +/-12 | 2,509 | Ai2 | Apache 2.0 |
| 205 | GPT-5-Nano (high)OpenAI | 1384.00 | +/-15 | 1,681 | OpenAI | Proprietary |
| 206 | Gemini 2.5 Flash-Lite (thinking)Google Deep Mind | 1383.00 | +/-8 | 5,998 | Google Deep Mind | Proprietary |
| 207 | qwen-plus-0125Alibaba | 1380.00 | +/-18 | 893 | Alibaba | Proprietary |
| 208 | llama-3.1-405b-instruct-bf16Meta | 1376.00 | +/-8 | 6,249 | Meta | Llama 3.1 Community |
| 209 | deepseek-v2.5-1210DeepSeek | 1375.00 | +/-17 | 1,079 | DeepSeek | DeepSeek |
| 210 | GPT-4.1 nanoOpenAI | 1374.00 | +/-19 | 807 | OpenAI | Proprietary |
| 211 | Llama 4 Maverick InstructFacebook AI研究实验室 | 1373.00 | +/-7 | 6,985 | Facebook AI研究实验室 | Llama 4 |
| 212 | hunyuan-turbo-0110Tencent | 1371.00 | +/-30 | 299 | Tencent | Proprietary |
| 213 | step-2-16k-exp-202412StepFun | 1371.00 | +/-20 | 737 | StepFun | Proprietary |
| 214 | GPT OSS 20BOpenAI | 1370.00 | +/-13 | 2,169 | OpenAI | Apache 2.0 |
| 215 | athene-v2-chat NexusFlow | 1369.00 | +/-9 | 4,019 | — | — |
| 216 | GPT-4oOpenAI | 1369.00 | +/-6 | 19,526 | OpenAI | Proprietary |
| 217 | yi-lightning Proprietary | 1369.00 | +/-10 | 4,316 | — | — |
| 218 | llama-3.1-405b-instruct-fp8Meta | 1368.00 | +/-7 | 9,714 | Meta | Llama 3.1 Community |
| 219 | DeepSeek V2.5DeepSeek-AI | 1368.00 | +/-9 | 4,252 | DeepSeek-AI | DeepSeek |
| 220 | mercury InceptionAI | 1368.00 | +/-29 | 393 | AI | Proprietary |
| 221 | hunyuan-large-2025-02-10Tencent | 1367.00 | +/-25 | 519 | Tencent | Proprietary |
| 222 | Gemini 2.0 Flash ExperimentalDeepMind | 1365.00 | +/-7 | 6,990 | DeepMind | Proprietary |
| 223 | olmo-3-32b-thinkAi2 | 1363.00 | +/-18 | 1,054 | Ai2 | Apache 2.0 |
| 224 | llama-3.3-nemotron-49b-super-v1Nvidia | 1363.00 | +/-31 | 286 | Nvidia | Nvidia |
| 225 | nvidia-nemotron-3-nano-30b-a3b-bf16Nvidia | 1363.00 | +/-10 | 3,280 | Nvidia | NVIDIA Open Model |
| 226 | Llama 4 Scout InstructFacebook AI研究实验室 | 1363.00 | +/-9 | 5,251 | Facebook AI研究实验室 | Llama |
| 227 | Mistral-Small-3.1-24B-Instruct-2503MistralAI | 1362.00 | +/-8 | 6,136 | MistralAI | Apache 2.0 |
| 228 | GPT-4oOpenAI | 1360.00 | +/-8 | 7,318 | OpenAI | Proprietary |
| 229 | 1358.00 | +/-7 | 10,368 | xAI | Proprietary | |
| 230 | Gemma 3 - 27B (IT)Google Deep Mind | 1358.00 | +/-7 | 8,067 | Google Deep Mind | Gemma |
| 231 | qwen2.5-plus-1127Alibaba | 1357.00 | +/-14 | 1,553 | Alibaba | Proprietary |
| 232 | Gemini 1.5 ProGoogle Deep Mind | 1356.00 | +/-7 | 9,175 | Google Deep Mind | Proprietary |
| 233 | hunyuan-large-visionTencent | 1356.00 | +/-19 | 963 | Tencent | Proprietary |
| 234 | Qwen2.5-VL-72B-Instruct阿里巴巴 | 1356.00 | +/-8 | 6,688 | 阿里巴巴 | Qwen |
| 235 | Claude3-OpusAnthropic | 1355.00 | +/-6 | 33,748 | Anthropic | Proprietary |
| 236 | mistral-large-2407Mistral | 1354.00 | +/-8 | 7,589 | Mistral | Mistral Research |
| 237 | step-1o-turbo-202506StepFun | 1353.00 | +/-15 | 1,505 | StepFun | Proprietary |
| 238 | qwen-max-0919Alibaba | 1353.00 | +/-11 | 2,756 | Alibaba | Qwen |
| 239 | granite-4.1-8bIBM | 1353.00 | +/-20 | 1,063 | IBM | Apache 2.0 |
| 240 | glm-4-plusZhipu AI | 1352.00 | +/-9 | 4,449 | Zhipu AI | Proprietary |
| 241 | athene-70b-0725 CC-BY-NC-4.0 | 1350.00 | +/-11 | 3,122 | — | — |
| 242 | GPT-4o miniOpenAI | 1349.00 | +/-7 | 10,926 | OpenAI | Proprietary |
| 243 | gpt-4-turbo-2024-04-09OpenAI | 1347.00 | +/-7 | 17,104 | OpenAI | Proprietary |
| 244 | Gemini 1.5 ProGoogle Deep Mind | 1347.00 | +/-8 | 12,747 | Google Deep Mind | Proprietary |
| 245 | mistral-large-2411Mistral | 1346.00 | +/-9 | 4,212 | Mistral | MRL |
| 246 | Llama3.3-70B-InstructFacebook AI研究实验室 | 1346.00 | +/-7 | 8,745 | Facebook AI研究实验室 | Llama-3.3 |
| 247 | amazon-nova-pro-v1.0Amazon | 1343.00 | +/-9 | 3,853 | Amazon | Proprietary |
| 248 | Gemini 2.0 Flash-LiteDeepMind | 1343.00 | +/-10 | 3,474 | DeepMind | Proprietary |
| 249 | deepseek-coder-v2DeepSeek | 1342.00 | +/-12 | 2,671 | DeepSeek | DeepSeek License |
| 250 | Qwen2.5-Coder-32B-Instruct阿里巴巴 | 1342.00 | +/-19 | 873 | 阿里巴巴 | Apache 2.0 |
| 251 | GPT-4OpenAI | 1340.00 | +/-7 | 15,605 | OpenAI | Proprietary |
| 252 | olmo-3.1-32b-thinkAi2 | 1338.00 | +/-15 | 1,569 | Ai2 | Apache 2.0 |
| 253 | gemini-advanced-0514Google | 1337.00 | +/-9 | 8,138 | Proprietary | |
| 254 | 1335.00 | +/-7 | 8,652 | xAI | Proprietary | |
| 255 | Llama3.1-70B-InstructFacebook AI研究实验室 | 1333.00 | +/-7 | 9,389 | Facebook AI研究实验室 | Llama 3.1 Community |
| 256 | hunyuan-standard-2025-02-10Tencent | 1332.00 | +/-24 | 549 | Tencent | Proprietary |
| 257 | GPT-4OpenAI | 1332.00 | +/-8 | 15,289 | OpenAI | Proprietary |
| 258 | glm-4-plus-0111Zhipu | 1330.00 | +/-18 | 894 | Zhipu | Proprietary |
| 259 | GPT-4OpenAI | 1329.00 | +/-9 | 8,306 | OpenAI | Proprietary |
| 260 | ibm-granite-h-smallIBM | 1329.00 | +/-17 | 1,267 | IBM | Apache 2.0 |
| 261 | Llama3.1-70B-InstructFacebook AI研究实验室 | 1328.00 | +/-15 | 1,312 | Facebook AI研究实验室 | Llama 3.1 |
| 262 | Claude3-SonnetAnthropic | 1318.00 | +/-7 | 18,888 | Anthropic | Proprietary |
| 263 | Gemma 3 - 12B (IT)Google Deep Mind | 1316.00 | +/-23 | 543 | Google Deep Mind | Gemma |
| 264 | gemini-1.5-flash-002Google | 1316.00 | +/-8 | 5,892 | Proprietary | |
| 265 | reka-core-20240904 Proprietary | 1315.00 | +/-15 | 1,216 | — | — |
| 266 | GPT-4OpenAI | 1314.00 | +/-8 | 13,719 | OpenAI | Proprietary |
| 267 | Mistral Small 24B Instruct 2501MistralAI | 1312.00 | +/-12 | 2,083 | MistralAI | Apache 2.0 |
| 268 | jamba-1.5-large Jamba Open | 1312.00 | +/-15 | 1,440 | — | — |
| 269 | llama-3.1-nemotron-51b-instructNvidia | 1312.00 | +/-22 | 665 | Nvidia | Llama 3.1 |
| 270 | gemini-1.5-flash-001Google | 1309.00 | +/-8 | 10,680 | Proprietary | |
| 271 | GLM4智谱AI | 1309.00 | +/-14 | 1,718 | 智谱AI | Proprietary |
| 272 | nemotron-4-340b-instructNvidia | 1308.00 | +/-11 | 3,254 | Nvidia | NVIDIA Open Model |
| 273 | llama-3.1-tulu-3-70bAi2 | 1308.00 | +/-24 | 450 | Ai2 | Llama 3.1 |
| 274 | Gemma-3n-E4BGoogle Deep Mind | 1308.00 | +/-10 | 3,526 | Google Deep Mind | Gemma |
| 275 | Phi-4-reasoningMicrosoft Azure | 1306.00 | +/-10 | 3,305 | Microsoft Azure | MIT |
| 276 | Llama3-70B-InstructFacebook AI研究实验室 | 1306.00 | +/-7 | 28,126 | Facebook AI研究实验室 | Llama 3 Community |
| 277 | amazon-nova-lite-v1.0Amazon | 1306.00 | +/-10 | 3,060 | Amazon | Proprietary |
| 278 | gemma-2-27b-itGoogle | 1305.00 | +/-6 | 12,088 | Gemma license | |
| 279 | Claude3-HaikuAnthropic | 1301.00 | +/-7 | 20,898 | Anthropic | Proprietary |
| 280 | hunyuan-standard-256kTencent | 1301.00 | +/-24 | 497 | Tencent | Proprietary |
| 281 | Qwen2-72B-Instruct阿里巴巴 | 1297.00 | +/-9 | 6,249 | 阿里巴巴 | Qianwen LICENSE |
| 282 | mistral-large-2402Mistral | 1295.00 | +/-9 | 10,418 | Mistral | Proprietary |
| 283 | C4AI Aya Vision 32BCohereAI | 1292.00 | +/-9 | 4,685 | CohereAI | CC-BY-NC-4.0 |
| 284 | reka-flash-20240904 Proprietary | 1291.00 | +/-15 | 1,207 | — | — |
| 285 | amazon-nova-micro-v1.0Amazon | 1288.00 | +/-10 | 2,981 | Amazon | Proprietary |
| 286 | Llama3.1-8B-InstructFacebook AI研究实验室 | 1288.00 | +/-26 | 478 | Facebook AI研究实验室 | Apache 2.0 |
| 287 | command-r-08-2024Cohere | 1281.00 | +/-13 | 1,783 | Cohere | CC-BY-NC-4.0 |
| 288 | command-r-plus-08-2024Cohere | 1280.00 | +/-14 | 1,675 | Cohere | CC-BY-NC-4.0 |
| 289 | olmo-2-0325-32b-instructAi2 | 1280.00 | +/-27 | 427 | Ai2 | Apache-2.0 |
| 290 | Qwen1.5-110B-Chat阿里巴巴 | 1279.00 | +/-10 | 4,763 | 阿里巴巴 | Qianwen LICENSE |
| 291 | reka-flash-21b-20240226-online Proprietary | 1277.00 | +/-13 | 2,879 | — | — |
| 292 | Mixtral-8x22B-Instruct-v0.1MistralAI | 1277.00 | +/-9 | 8,780 | MistralAI | Apache 2.0 |
| 293 | Qwen1.5-72B-Chat阿里巴巴 | 1275.00 | +/-10 | 6,370 | 阿里巴巴 | Qianwen LICENSE |
| 294 | ministral-8b-2410Mistral | 1275.00 | +/-19 | 838 | Mistral | MRL |
| 295 | gpt-3.5-turbo-0125OpenAI | 1274.00 | +/-8 | 11,130 | OpenAI | Proprietary |
| 296 | Gemma 3 - 4B (IT)Google Deep Mind | 1274.00 | +/-24 | 605 | Google Deep Mind | Gemma |
| 297 | C4AI Command R+CohereAI | 1272.00 | +/-8 | 13,937 | CohereAI | CC-BY-NC-4.0 |
| 298 | gemini-1.5-flash-8b-001Google | 1272.00 | +/-8 | 6,069 | Proprietary | |
| 299 | gemma-2-9b-it-simpo MIT | 1271.00 | +/-15 | 1,471 | — | — |
| 300 | gemma-2-9b-itGoogle | 1270.00 | +/-7 | 8,921 | Gemma license | |
| 301 | reka-flash-21b-20240226 Proprietary | 1267.00 | +/-11 | 4,748 | — | — |
| 302 | jamba-1.5-mini Jamba Open | 1265.00 | +/-15 | 1,352 | — | — |
| 303 | mistral-mediumMistral | 1262.00 | +/-10 | 5,149 | Mistral | Proprietary |
| 304 | gpt-3.5-turbo-1106OpenAI | 1262.00 | +/-16 | 2,121 | OpenAI | Proprietary |
| 305 | qwen1.5-32b-chatAlibaba | 1262.00 | +/-11 | 3,930 | Alibaba | Qianwen LICENSE |
| 306 | Llama3.1-8B-InstructFacebook AI研究实验室 | 1260.00 | +/-7 | 8,582 | Facebook AI研究实验室 | Llama 3.1 Community |
| 307 | C4AI Aya Vision 8BCohereAI | 1255.00 | +/-15 | 1,567 | CohereAI | CC-BY-NC-4.0 |
| 308 | llama-3.1-tulu-3-8bAi2 | 1253.00 | +/-25 | 476 | Ai2 | Llama 3.1 |
| 309 | Llama3-8B-InstructFacebook AI研究实验室 | 1253.00 | +/-8 | 18,374 | Facebook AI研究实验室 | Llama 3 Community |
| 310 | dbrx-instruct-preview DBRX LICENSE | 1251.00 | +/-11 | 5,502 | — | — |
| 311 | granite-3.1-2b-instructIBM | 1249.00 | +/-25 | 508 | IBM | Apache 2.0 |
| 312 | Gemini-proDeepMind | 1248.00 | +/-24 | 678 | DeepMind | Proprietary |
| 313 | InternLM2-Base-20B上海人工智能实验室 | 1248.00 | +/-14 | 1,684 | 上海人工智能实验室 | — |
| 314 | Yi-1.5-34B零一万物 | 1248.00 | +/-10 | 3,841 | 零一万物 | — |
| 315 | zephyr-orpo-141b-A35b-v0.1 Apache 2.0 | 1245.00 | +/-21 | 831 | — | — |
| 316 | command-rCohere | 1243.00 | +/-9 | 9,645 | Cohere | CC-BY-NC-4.0 |
| 317 | granite-3.0-8b-instructIBM | 1240.00 | +/-18 | 1,108 | IBM | Apache 2.0 |
| 318 | Qwen1.5-14B-Chat阿里巴巴 | 1239.00 | +/-13 | 3,208 | 阿里巴巴 | Qianwen LICENSE |
| 319 | gemini-pro-dev-apiGoogle | 1239.00 | +/-14 | 2,681 | Proprietary | |
| 320 | mixtral-8x7b-instruct-v0.1Mistral | 1239.00 | +/-8 | 11,784 | Mistral | Apache 2.0 |
| 321 | starling-lm-7b-beta Apache-2.0 | 1235.00 | +/-13 | 2,948 | — | — |
| 322 | Phi-3-medium 14B-previewMicrosoft Azure | 1231.00 | +/-10 | 3,973 | Microsoft Azure | MIT |
| 323 | openchat-3.5-0106 Apache-2.0 | 1229.00 | +/-14 | 2,005 | — | — |
| 324 | snowflake-arctic-instruct Apache 2.0 | 1224.00 | +/-11 | 5,734 | — | — |
| 325 | DeepSeek LLM 67B ChatDeepSeek-AI | 1218.00 | +/-24 | 649 | DeepSeek-AI | DeepSeek License |
| 326 | Gemma 1.1-7B-ITGoogle Research | 1216.00 | +/-10 | 4,332 | Google Research | Gemma license |
| 327 | LLaMA2 70BFacebook AI研究实验室 | 1214.00 | +/-21 | 805 | Facebook AI研究实验室 | — |
| 328 | Qwen1.5-7B-Chat阿里巴巴 | 1209.00 | +/-21 | 772 | 阿里巴巴 | Qianwen LICENSE |
| 329 | Qwen3-VL-2B阿里巴巴 | 1209.00 | +/-17 | 1,134 | 阿里巴巴 | Apache 2.0 |
| 330 | starling-lm-7b-alpha CC-BY-NC-4.0 | 1207.00 | +/-16 | 1,397 | — | — |
| 331 | Yi-34B零一万物 | 1205.00 | +/-13 | 2,345 | 零一万物 | — |
| 332 | Phi-3-small 7BMicrosoft Azure | 1204.00 | +/-12 | 3,219 | Microsoft Azure | MIT |
| 333 | openchat-3.5 Apache-2.0 | 1202.00 | +/-20 | 971 | — | — |
| 334 | Qwen-14B-Chat阿里巴巴 | 1197.00 | +/-24 | 599 | 阿里巴巴 | Qianwen LICENSE |
| 335 | Phi-3-mini 3.8BMicrosoft Azure | 1197.00 | +/-14 | 1,841 | Microsoft Azure | MIT |
| 336 | Vicuna 33BLM-SYS | 1193.00 | +/-13 | 2,866 | LM-SYS | — |
| 337 | WizardLM-70B-V1.0WizardLM Team | 1193.00 | +/-20 | 988 | WizardLM Team | Llama 2 Community |
| 338 | gemma-2-2b-itGoogle | 1192.00 | +/-8 | 7,298 | Gemma license | |
| 339 | Phi-3-mini 3.8BMicrosoft Azure | 1187.00 | +/-12 | 3,449 | Microsoft Azure | MIT |
| 340 | openhermes-2.5-mistral-7b Apache-2.0 | 1186.00 | +/-23 | 589 | — | — |
| 341 | Mistral-7B-Instruct-v0.2MistralAI | 1185.00 | +/-12 | 3,114 | MistralAI | Apache-2.0 |
| 342 | solar-10.7b-instruct-v1.0 CC-BY-NC-4.0 | 1183.00 | +/-27 | 482 | — | — |
| 343 | llama-2-70b-chatMeta | 1179.00 | +/-10 | 5,717 | Meta | Llama 2 Community |
| 344 | llama-3.2-3b-instructMeta | 1176.00 | +/-16 | 1,351 | Meta | Llama 3.2 |
| 345 | nous-hermes-2-mixtral-8x7b-dpo Apache-2.0 | 1175.00 | +/-23 | 575 | — | — |
| 346 | QwQ-32B-Preview阿里巴巴 | 1173.00 | +/-24 | 566 | 阿里巴巴 | Apache 2.0 |
| 347 | Gemma 1.1-2B-ITGoogle Research | 1171.00 | +/-14 | 1,963 | Google Research | Gemma license |
| 348 | mpt-30b-chat CC-BY-NC-SA-4.0 | 1167.00 | +/-34 | 258 | — | — |
| 349 | Gemma 7B - ItGoogle Research | 1167.00 | +/-17 | 1,381 | Google Research | Gemma license |
| 350 | zephyr-7b-alpha MIT | 1166.00 | +/-39 | 201 | — | — |
| 351 | vicuna-13b Llama 2 Community | 1163.00 | +/-14 | 2,389 | — | — |
| 352 | Baichuan2-13B-Chat百川智能 | 1162.00 | +/-13 | 2,626 | 百川智能 | Llama 2 Community |
| 353 | smollm2-1.7b-instruct Apache 2.0 | 1160.00 | +/-33 | 352 | — | — |
| 354 | CodeLLaMA-34BFacebook AI研究实验室 | 1159.00 | +/-20 | 853 | Facebook AI研究实验室 | Llama 2 Community |
| 355 | Phi-3-mini 3.8BMicrosoft Azure | 1154.00 | +/-13 | 3,886 | Microsoft Azure | MIT |
| 356 | PaLM 2Google Research | 1153.00 | +/-21 | 917 | Google Research | Proprietary |
| 357 | zephyr-7b-beta MIT | 1152.00 | +/-18 | 1,250 | — | — |
| 358 | wizardlm-13bMicrosoft | 1151.00 | +/-22 | 735 | Microsoft | Llama 2 Community |
| 359 | llama-3.2-1b-instructMeta | 1149.00 | +/-16 | 1,346 | Meta | Llama 3.2 |
| 360 | llama2-70b-steerlm-chatNvidia | 1145.00 | +/-28 | 467 | Nvidia | Llama 2 Community |
| 361 | Mistral 7B InstructMistralAI | 1144.00 | +/-20 | 1,032 | MistralAI | Apache 2.0 |
| 362 | Gemma 2B - ItGoogle Research | 1136.00 | +/-22 | 742 | Google Research | Gemma license |
| 363 | vicuna-7b Llama 2 Community | 1131.00 | +/-23 | 726 | — | — |
| 364 | Qwen1.5-4B-Chat阿里巴巴 | 1131.00 | +/-17 | 1,283 | 阿里巴巴 | Qianwen LICENSE |
| 365 | stripedhyena-nous-7b Apache 2.0 | 1127.00 | +/-22 | 704 | — | — |
| 366 | guanaco-33b Non-commercial | 1113.00 | +/-36 | 263 | — | — |
| 367 | olmo-7b-instructAi2 | 1107.00 | +/-22 | 772 | Ai2 | Apache-2.0 |
| 368 | Baichuan2-7B-Chat百川智能 | 1103.00 | +/-14 | 1,956 | 百川智能 | Llama 2 Community |
| 369 | ChatGLM3-6B智谱AI | 1090.00 | +/-26 | 535 | 智谱AI | — |
| 370 | mpt-7b-chat CC-BY-NC-SA-4.0 | 1066.00 | +/-31 | 397 | — | — |
| 371 | Koala达摩院 | 1066.00 | +/-24 | 747 | 达摩院 | — |
| 372 | RWKV-4-Raven-14B Apache 2.0 | 1059.00 | +/-27 | 505 | — | — |
| 373 | oasst-pythia-12b Apache 2.0 | 1050.00 | +/-25 | 714 | — | — |
| 374 | ChatGLM-6B智谱AI | 1035.00 | +/-27 | 551 | 智谱AI | — |
| 375 | ChatGLM2-6B智谱AI | 1030.00 | +/-35 | 293 | 智谱AI | — |
| 376 | stablelm-tuned-alpha-7b CC-BY-NC-SA-4.0 | 1004.00 | +/-32 | 363 | — | — |
| 377 | alpaca-13b Non-commercial | 999.00 | +/-27 | 626 | — | — |
| 378 | dolly-v2-12b MIT | 962.00 | +/-34 | 396 | — | — |
| 379 | fastchat-t5-3b Apache 2.0 | 907.00 | +/-30 | 428 | — | — |
| 380 | LLaMA 13BFacebook AI研究实验室 | 882.00 | +/-39 | 304 | Facebook AI研究实验室 | Non-commercial |
Data is for reference only. Official sources are authoritative. Click model names to view DataLearner model profiles.
FAQ
What is LMArena Coding Arena?
LMArena Coding Arena is an anonymous evaluation track focused on coding ability. Users submit real programming tasks such as debugging, code generation, and algorithm implementation; two hidden model answers are shown side by side, and user votes are aggregated into an Elo leaderboard.
How is Coding Arena different from SWE-bench or HumanEval?
Static benchmarks use fixed test sets and automated scoring, which makes them reproducible but easier to over-optimize for. Coding Arena uses open-ended user tasks and human preference votes, so it better reflects practical coding experience. The two approaches are complementary.
How do China-developed models perform on coding tasks?
Models such as DeepSeek and Qwen rank competitively on coding leaderboards. They are especially relevant when open deployment, Chinese-language developer workflows, or cost control matter.
How can AI help with day-to-day programming?
Common workflows include code completion and generation, debugging, code review, unit test generation, and cross-language translation.
















