DataLearner logo

Artificial Analysis Intelligence Index

Artificial Analysis Intelligence Index aggregates multiple rigorous benchmarks to compare AI model intelligence across coding, reasoning, science, tool use, and agentic tasks.

Top Model

Claude Opus 5 (max)

Top Score

63

Model Count

256

Data version

2026年08月18日

Data source: Artificial Analysis

Origin:AllChina
Leaderboard snapshot month:

Ranking Table

RankModelIntelligence IndexOrganization
AnthropicClaude Opus 5 (max)Anthropic63Anthropic
AnthropicClaude Opus 5 (xhigh)Anthropic63Anthropic
AnthropicClaude Fable 5Anthropic62Anthropic
4AnthropicClaude Opus 5 (high)Anthropic61Anthropic
5OpenAIGPT-5.6 Sol (max)OpenAI61OpenAI
6xAIGrok 4.6 (high)xAI61xAI
7Moonshot AIKimi K3 (max)Moonshot AI60Moonshot AI
8OpenAIGPT-5.6 Sol (xhigh)OpenAI59OpenAI
9AnthropicClaude Opus 5 (medium)Anthropic59Anthropic
10阿里巴巴Qwen3.8-Max阿里巴巴58阿里巴巴
11AlibabaQwen3.8 2.4T A95BAlibaba58Alibaba
12OpenAIGPT-5.6 Sol (high)OpenAI57OpenAI
13MetaMuse Spark 1.2 (xhigh)Meta57Meta
14OpenAIGPT-5.6 Terra (max)OpenAI57OpenAI
15Google Deep MindGemini 3.7 Flash (high)Google Deep Mind56Google Deep Mind
16xAIGrok 4.5 (high)xAI56xAI
17OpenAIGPT-5.6 Sol (medium)OpenAI56OpenAI
18AnthropicClaude Sonnet 5 (max)Anthropic55Anthropic
19Google Deep MindGemini 3.7 Flash (medium)Google Deep Mind53Google Deep Mind
20DeepSeek-AIDeepSeek-V4-Pro (max)DeepSeek-AI53DeepSeek-AI
21OpenAIGPT-5.6 Terra (xhigh)OpenAI53OpenAI
22智谱AIGLM-5.2 (max)智谱AI53智谱AI
23AnthropicClaude Opus 5 (low)Anthropic52Anthropic
24OpenAIGPT-5.6 Luna (max)OpenAI52OpenAI
25阿里巴巴Qwen3.8-27B阿里巴巴52阿里巴巴
26DeepSeek-AIDeepSeek-V4-Flash (max)DeepSeek-AI52DeepSeek-AI
27Google Deep MindGemini 3.6 FlashGoogle Deep Mind52Google Deep Mind
28Google Deep MindGemini 3.7 Flash (low)Google Deep Mind51Google Deep Mind
29OpenAIGPT-5.6 Sol (low)OpenAI51OpenAI
30OpenAIGPT-5.6 Terra (high)OpenAI50OpenAI
31OpenAIGPT-5.6 Luna (xhigh)OpenAI50OpenAI
32Moonshot AIKimi K3 (low)Moonshot AI48Moonshot AI
33Google Deep MindGemini 3.1 Pro PreviewGoogle Deep Mind48Google Deep Mind
34Motif 3Motif Technologies47Motif Technologies
35OpenAIGPT-5.6 Luna (high)OpenAI47OpenAI
36OpenAIGPT-5.6 Terra (medium)OpenAI47OpenAI
37Google Deep MindGemini 3.5 Flash (medium)Google Deep Mind47Google Deep Mind
38OpenAIGPT-5.3 Codex (xhigh)OpenAI46OpenAI
39MiniMaxAIMiniMax M3MiniMaxAI45MiniMaxAI
40Motif 3 (Beta)Motif Technologies45Motif Technologies
41DeepSeek-AIDeepSeek-V4-Pro (max)DeepSeek-AI45DeepSeek-AI
42DeepSeek-AIDeepSeek-V4-Pro (high)DeepSeek-AI44DeepSeek-AI
43Moonshot AIKimi K2.7 CodeMoonshot AI43Moonshot AI
44MiMo-V2.5-ProXiaomi43Xiaomi
45AnthropicClaude Sonnet 5 (non-reasoning)Anthropic43Anthropic
46InklingThinking Machines Lab42Thinking Machines Lab
47腾讯AI实验室Hy3腾讯AI实验室42腾讯AI实验室
48OpenAIGPT-5.6 Sol (non-reasoning)OpenAI42OpenAI
49Nex-N2-ProNex AGI42Nex AGI
50Solar Pro 4Upstage42Upstage
51OpenAIGPT-5.6 Terra (low)OpenAI41OpenAI
52Inkling SmallThinking Machines41Thinking Machines
53JT-4.1 Flash 236B A21BChina Mobile40China Mobile
54Agnes 2.5 Pro AlphaSapiens AI40Sapiens AI
55阿里巴巴Qwen3.7-Plus阿里巴巴39阿里巴巴
56OpenAIGPT-5.6 Luna (medium)OpenAI39OpenAI
57NVIDIANemotron 3 UltraNVIDIA38NVIDIA
58MiMo-V2.5Xiaomi38Xiaomi
59Ling 3.0 FlashInclusionAI38InclusionAI
60阿里巴巴Qwen3.6-27B阿里巴巴38阿里巴巴
61Google Deep MindGemini 3.5 Flash-LiteGoogle Deep Mind37Google Deep Mind
62Solar Open2 250BUpstage37Upstage
63亚马逊Nova 2 Omni(Preview)亚马逊37亚马逊
64xAIGrok 4.3 Beta (medium)xAI37xAI
65xAIGrok 4.3 Beta (low)xAI36xAI
66MiMo-V2-OmniXiaomi36Xiaomi
67Google Deep MindGemini 3.5 Flash (minimal)Google Deep Mind36Google Deep Mind
68AnthropicClaude Sonnet 4.6 (Non-reasoning, Low Effort)Anthropic35Anthropic
69Facebook AI研究实验室Muse Glimmer-30B (high)Facebook AI研究实验室35Facebook AI研究实验室
70阿里巴巴Qwen2-57B-A14B阿里巴巴35阿里巴巴
71智谱AIGLM-5.2智谱AI35智谱AI
72OpenAIGPT-5.6 Terra (non-reasoning)OpenAI35OpenAI
73阿里巴巴Qwen3.5-397B-A17B阿里巴巴34阿里巴巴
74DeepMindGemini 2.0 Flash ExperimentalDeepMind34DeepMind
75LongCat 2.0LongCat34LongCat
76OpenAIGPT-5.6 Luna (low)OpenAI34OpenAI
77KAT-Coder-Pro V2KwaiKAT34KwaiKAT
78阿里巴巴Qwen3.5-122B-A10B阿里巴巴33阿里巴巴
79阿里巴巴Qwen3.5-397B-A17B阿里巴巴33阿里巴巴
80阿里巴巴Qwen3.6-35B-A3B阿里巴巴32阿里巴巴
81DeepSeek-AIDeepSeek-V4-ProDeepSeek-AI32DeepSeek-AI
82Ring-2.6-1TInclusionAI32InclusionAI
83G9v3-39A5BAI9Stars32AI9Stars
84阿里巴巴Qwen3.5-Omni-Plus阿里巴巴31阿里巴巴
85阿里巴巴Qwen3.6-27B阿里巴巴31阿里巴巴
86OpenAIOpenAI o3OpenAI31OpenAI
87K-EXAONE 2.0LG AI Research31LG AI Research
88StepFunAIStep 3.7 FlashStepFunAI31StepFunAI
89MistralAIMistral Medium 3.5MistralAI30MistralAI
90AnthropicHaiku 4.5Anthropic30Anthropic
91DeepMindGemma 4 31BDeepMind30DeepMind
92DeepSeek-AIDeepSeek-V4-FlashDeepSeek-AI29DeepSeek-AI
93OpenAIGPT-5.5 InstantOpenAI29OpenAI
94JT-35B-FlashChina Mobile29China Mobile
95MiMo-V2.5-ProXiaomi28Xiaomi
96阿里巴巴Qwen3.5-122B-A10B阿里巴巴28阿里巴巴
97OpenAIGPT-5.6 Luna (non-reasoning)OpenAI27OpenAI
98ByteDance SeedDoubao Seed CodeByteDance Seed26ByteDance Seed
99DeepMindGemma 4 26B A4BDeepMind26DeepMind
100NVIDIANemotron 3 SuperNVIDIA26NVIDIA
101MiMo-V2-FlashXiaomi25Xiaomi
102xAIGrok 4.3 Beta (non-reasoning)xAI25xAI
103阿里巴巴Qwen3.6-35B-A3B阿里巴巴25阿里巴巴
104Ling 3.0 TinyInclusionAI25InclusionAI
105阿里巴巴Qwen3.5-35B-A3B阿里巴巴24阿里巴巴
106AnthropicHaiku 4.5Anthropic24Anthropic
107OpenAIGPT OSS 120B (high)OpenAI24OpenAI
108NVIDIANemotron 3.5 LightningNVIDIA24NVIDIA
109CohereAIC4AI Command A (202503)CohereAI23CohereAI
110K-EXAONELG AI Research22LG AI Research
111百度ERNIE 5.0 Thinking Preview百度22百度
112DeepMindGemma 4 31BDeepMind22DeepMind
113GoogleGemma 4 12BGoogle22Google
114亚马逊Nova 2 Pro(Preview) (medium)亚马逊22亚马逊
115Mercury 2Inception22Inception
116阿里巴巴Qwen3.5-9B阿里巴巴22阿里巴巴
117阿里巴巴Qwen3-Coder-Next阿里巴巴21阿里巴巴
118亚马逊Nova 2 Omni(Preview) (medium)亚马逊21亚马逊
119Apriel-v1.6-15B-ThinkerServiceNow21ServiceNow
120亚马逊Nova 2 Lite (high)亚马逊21亚马逊
121阿里巴巴Qwen3.5-9B阿里巴巴21阿里巴巴
122EXAONE 4.5 33BLG AI Research21LG AI Research
123DeepMindGemma 4 26B A4BDeepMind20DeepMind
124AlibabaQwen3.5 4BAlibaba20Alibaba
125CohereNorth Mini CodeCohere20Cohere
126亚马逊Nova 2 Pro(Preview) (low)亚马逊20亚马逊
127MistralMistral Small 4Mistral20Mistral
128MistralDevstral 2Mistral19Mistral
129亚马逊Nova 2 Lite (medium)亚马逊19亚马逊
130阿里巴巴Qwen3.5-Omni-Flash阿里巴巴19阿里巴巴
131JT-MINIChina Mobile19China Mobile
132Trinity Large ThinkingArcee AI19Arcee AI
133HyperNova 60B 2605Multiverse Computing18Multiverse Computing
134MistralMagistral Medium 1.2Mistral18Mistral
135亚马逊Nova 2 Lite (low)亚马逊18亚马逊
136NVIDIANemotron Cascade 2 30B A3BNVIDIA18NVIDIA
137MistralDevstral Small 2Mistral18Mistral
138K2 Think V2MBZUAI17MBZUAI
139LongCat Flash LiteLongCat17LongCat
140HyperCLOVA X SEED Think (32B)Naver17Naver
141K-EXAONELG AI Research17LG AI Research
142阿里巴巴Qwen3-Next阿里巴巴17阿里巴巴
143亚马逊Nova 2 Omni(Preview) (low)亚马逊17亚马逊
144Mi:dm K 2.5 ProKorea Telecom17Korea Telecom
145G9v3-3BAI9Stars16AI9Stars
146AlibabaQwen3.5 4BAlibaba16Alibaba
147MistralAIMistral Large 3MistralAI16MistralAI
148INTELLECT-3Prime Intellect16Prime Intellect
149Solar Open 100BUpstage15Upstage
150OpenAIGPT OSS 20B (high)OpenAI15OpenAI
151阿里巴巴Qwen3-Omni-30B-A3B阿里巴巴15阿里巴巴
152OpenAIGPT OSS 120B (low)OpenAI15OpenAI
153NVIDIANemotron 3 NanoNVIDIA15NVIDIA
154Solar Pro 3Upstage14Upstage
155Facebook AI研究实验室Llama 4 MaverickFacebook AI研究实验室14Facebook AI研究实验室
156亚马逊Nova 2 Pro(Preview)亚马逊14亚马逊
157OpenAIGPT OSS 20B (low)OpenAI14OpenAI
158K2-V2 (high)MBZUAI14MBZUAI
159阿里巴巴Qwen3-Next阿里巴巴14阿里巴巴
160GoogleDiffusionGemma 26B A4BGoogle13Google
161GoogleGemma 4 12B (Non-reasoning)Google13Google
162Motif-2-12.7BMotif Technologies13Motif Technologies
163AmazonNova PremierAmazon13Amazon
164K2-V2 (medium)MBZUAI12MBZUAI
165MetaLlama Nemotron Super 49B v1.5Meta12Meta
166Celeris-1Celeris12Celeris
167MistralMistral Small 4Mistral12Mistral
168Tri-21B-ThinkTrillion Labs12Trillion Labs
169DeepMindGemma 4 E4BDeepMind12DeepMind
170OpenBMBMiniCPM5-1BOpenBMB12OpenBMB
171Sarvam 105B (high)Sarvam12Sarvam
172亚马逊Nova 2 Lite亚马逊12亚马逊
173OpenBMBMiniCPM5-1BOpenBMB12OpenBMB
174MistralMagistral Small 1.2Mistral11Mistral
175MistralAIMinistral 3 14BMistralAI11MistralAI
176Nanbeige4.1-3BNanbeige11Nanbeige
177EXAONE 4.0 32BLG AI Research11LG AI Research
178亚马逊Nova 2 Omni(Preview)亚马逊10亚马逊
179Facebook AI研究实验室Llama 4 ScoutFacebook AI研究实验室10Facebook AI研究实验室
180Hermes 4 70BNous Research10Nous Research
181Falcon-H1R-7BTII UAE10TII UAE
182DeepMindGemma 4 E2BDeepMind10DeepMind
183阿里巴巴Qwen3-Omni-30B-A3B阿里巴巴10阿里巴巴
184StepFunStep3 VL 10BStepFun9StepFun
185Facebook AI研究实验室Llama3.3-70B-InstructFacebook AI研究实验室9Facebook AI研究实验室
186MistralAIMinistral 3 8BMistralAI9MistralAI
187NVIDIALlama Nemotron UltraNVIDIA9NVIDIA
188百度ERNIE-4.5-300B-A47B百度9百度
189Hermes 4 405BNous Research9Nous Research
190NVIDIANVIDIA Nemotron Nano 12B v2 VLNVIDIA9NVIDIA
191DeepMindGemma 4 E4BDeepMind9DeepMind
192Granite 4.1 30BIBM9IBM
193NVIDIANVIDIA Nemotron Nano 9B V2NVIDIA9NVIDIA
194Hermes 4 405BNous Research9Nous Research
195NVIDIANemotron 3 Nano 4BNVIDIA9NVIDIA
196MetaLlama Nemotron Super 49B v1.5Meta9Meta
197K2-V2 (low)MBZUAI8MBZUAI
198KimiKimi Linear 48B A3B InstructKimi8Kimi
199Facebook AI研究实验室Llama3.1-405BFacebook AI研究实验室8Facebook AI研究实验室
200LFM2.5-8B-A1BLiquid AI8Liquid AI
201Ring-flash-2.0InclusionAI8InclusionAI
202Olmo 3.1 32B ThinkAI28AI2
203CohereAIC4AI Command A (202503)CohereAI7CohereAI
204AlibabaQwen3.5 2BAlibaba7Alibaba
205NVIDIALlama 3.1 Nemotron 70BNVIDIA7NVIDIA
206NVIDIANemotron 3 NanoNVIDIA7NVIDIA
207NVIDIANVIDIA Nemotron Nano 9B V2NVIDIA7NVIDIA
208MistralMinistral 3 3BMistral7Mistral
209Hermes 4 70BNous Research7Nous Research
210Granite 4.1 8BIBM6IBM
211Sarvam 30B (high)Sarvam6Sarvam
212Olmo 3.1 32B InstructAI26AI2
213DeepMindGemma 4 E2BDeepMind6DeepMind
214PerplexityR1 1776Perplexity6Perplexity
215Facebook AI研究实验室Llama 3.2-Vision-90BFacebook AI研究实验室6Facebook AI研究实验室
216Microsoft AzurePhi-4-mini-instruct (3.8B)Microsoft Azure6Microsoft Azure
217EXAONE 4.0 32BLG AI Research6LG AI Research
218AlibabaQwen3.5 2BAlibaba5Alibaba
219AlibabaQwen3.5 0.8BAlibaba5Alibaba
220DeepHermes 3 - Mistral 24BNous Research5Nous Research
221Jamba 1.7 LargeAI21 Labs5AI21 Labs
222Granite 4.0 H SmallIBM5IBM
223阿里巴巴Qwen3-Omni-30B-A3B阿里巴巴5阿里巴巴
224LFM2 24B A2BLiquid AI5Liquid AI
225Microsoft AzurePhi-4-reasoningMicrosoft Azure5Microsoft Azure
226亚马逊Amazon Nova Micro亚马逊4亚马逊
227Granite 4.1 3BIBM4IBM
228NVIDIANVIDIA Nemotron Nano 12B v2 VLNVIDIA4NVIDIA
229Microsoft AzurePhi-4-multimodal-instruct Microsoft Azure4Microsoft Azure
230MiniCPM-V 4.6 1.3BOpenBMB4OpenBMB
231Jamba Reasoning 3BAI21 Labs4AI21 Labs
232Google Deep MindGemini 3.0 FlashGoogle Deep Mind4Google Deep Mind
233Olmo 3 7B ThinkAI24AI2
234Molmo 7B-DAllen Institute for AI3Allen Institute for AI
235Facebook AI研究实验室Llama 3.2-Vision-11BFacebook AI研究实验室3Facebook AI研究实验室
236AlibabaQwen3.5 0.8BAlibaba3Alibaba
237Exaone 4.0 1.2BLG AI Research3LG AI Research
238Olmo 3 7BAI22AI2
239Exaone 4.0 1.2BLG AI Research2LG AI Research
240LFM2.5-1.2B-ThinkingLiquid AI2Liquid AI
241Jamba 1.7 MiniAI21 Labs2AI21 Labs
242LFM2 2.6BLiquid AI2Liquid AI
243LFM2.5-1.2B-InstructLiquid AI2Liquid AI
244Granite 4.0 H 1BIBM2IBM
245Google Deep MindGemma 3-270MGoogle Deep Mind2Google Deep Mind
246Apertus 70B InstructSwiss AI2Swiss AI
247Granite 4.0 MicroIBM2IBM
248DeepHermes 3 - Llama-3.1 8BNous Research2Nous Research
249Granite 4.0 1BIBM2IBM
250Molmo2-8BAI22AI2
251LFM2 8B A1BLiquid AI1Liquid AI
252LFM2.5-VL-1.6BLiquid AI1Liquid AI
253Granite 4.0 350MIBM1IBM
254CohereTiny Aya GlobalCohere1Cohere
255Apertus 8B InstructSwiss AI1Swiss AI
256Granite 4.0 H 350MIBM1IBM

Data is for reference only. Official sources are authoritative. Click model names to view DataLearner model profiles.

Benchmark Components (Intelligence Index v4.0)

The Intelligence Index aggregates 10 rigorous benchmarks to provide a holistic measure of AI capabilities, preventing narrow specialization.

GDPval-AA
Agentic real-world tasks
τ²-Bench
Agentic tool use
Terminal-Bench
Agentic coding
SciCode
Coding proficiency
AA-LCR
Long context reasoning
AA-Omniscience
Knowledge & hallucination
IFBench
Instruction following
Humanity's Last Exam
Reasoning & knowledge
GPQA Diamond
Scientific reasoning
CritPt
Physics reasoning

FAQ

What is the Artificial Analysis Intelligence Index?
The Artificial Analysis Intelligence Index v4.0 is a composite benchmark that aggregates performance across 10 challenging evaluations — spanning mathematics, science, coding, agentic tasks, and reasoning — to measure AI capabilities holistically. It is designed to prevent narrow specialization and provide a single score for tracking progress.
How is the Intelligence Index calculated?
The index aggregates scores from 10 benchmarks: GDPval-AA (agentic real-world tasks), τ²-Bench (tool use), Terminal-Bench Hard (agentic coding), SciCode (coding), AA-LCR (long context reasoning), AA-Omniscience (knowledge & hallucination), IFBench (instruction following), Humanity's Last Exam (reasoning), GPQA Diamond (scientific reasoning), and CritPt (physics). All tests are independently run by Artificial Analysis on standardized hardware.
How does this differ from LMArena?
LMArena rankings are based on crowdsourced user votes (Elo ratings from blind A/B tests), reflecting subjective human preferences. The Artificial Analysis Intelligence Index uses standardized automated benchmarks with objective scoring, measuring technical capabilities across specific domains. Both perspectives are valuable — LMArena captures real-world user experience, while AA Intelligence Index provides reproducible technical measurements.
Where can I find the original data?
The original leaderboard and detailed methodology are available at artificialanalysis.ai. The Intelligence Index methodology is documented at Intelligence Index page.