Skip to main content
ModelScale

Knowledge benchmark

Artificial Analysis Intelligence Index leaderboard

Every model the catalog carries a published Artificial Analysis Intelligence Index value for, ranked by that value.

CategoryKnowledge
MeasureAggregated model score
TasksCross-benchmark intelligence index
DifficultyDisplay-only external reference

A display-only intelligence index published by Artificial Analysis that aggregates provider-reported and benchmark-derived signals into a single model-level score.

Artificial Analysis Intelligence Index ranking

185 models with a published Artificial Analysis Intelligence Index value, ordered by that value, highest first. Models the source has not scored on this benchmark are not listed — they are unmeasured, not last.

Models ranked by their published Artificial Analysis Intelligence Index value
RankModelProviderAggregated model score
1GPT-5.6 SolOpenAI58.9
2GPT-5.6 TerraOpenAI55.0
3Claude Fable 5.1Anthropic53.4
4DeepSeek V4 Pro 0813DeepSeek53.2
5GPT-6 AstraOpenAI52.8
6GPT-5.6 LunaOpenAI51.2
7Claude Opus 5Anthropic50.7
8Gemini 3.5 FlashGoogle50.2
9Claude Fable 5Anthropic49.7
10Muse Spark 1.3Meta48.2
11Qwen3.8 Max PreviewAlibaba45.4
12GLM-5.3Z.AI44.9
13Grok 4.6xAI44.4
14Kimi K3Moonshot AI43.8
15Claude Opus 4.8Anthropic42.0
16GLM-5.3-FlashZ.AI41.9
17Hy3 PreviewTencent41.2
18Gemini 3.8 FlashGoogle41.2
19Claude Opus 4.7 (Adaptive)Anthropic40.7
20Qwen3.8-Flash-NextAlibaba39.9
21Muse Spark 1.2Meta39.8
22DeepSeek V4.1 FlashDeepSeek39.5
23Gemini 3.7 FlashGoogle39.4
24Grok 4.5xAI39.1
25GPT-5.4OpenAI39.0
26GPT-5.5OpenAI38.6
27Claude Sonnet 5Anthropic38.4
28Grok 4.3xAI37.6
29DeepSeek V4 Flash 0731DeepSeek34.5
30Gemini 3.6 FlashGoogle34.3
31Muse Spark 1.1Meta34.3
32GLM-5.2Z.AI34.0
33Qwen3.8-27BAlibaba33.9
34GPT-5.3 CodexOpenAI32.5
35Claude Opus 4.6 (Adaptive)Anthropic31.9
36Muse SparkMeta31.3
37Claude Opus 4.7Anthropic30.9
38GPT-5.2OpenAI30.4
39APApodex 1.1Apodex30.4
39APApodex 1.1 MiniApodex30.4
41Gemini 3.1 ProGoogle30.4
42Qwen3.7 MaxAlibaba29.9
43MiniMax M3MiniMax29.6
44Claude Opus 4.5 ThinkingAnthropic29.1
45MiMo-V2-ProXiaomi28.6
46GPT-5.2-CodexOpenAI28.5
47Qwen 3.6 Max (preview)Alibaba28.4
48Solar Pro 4Upstage28.1
49Gemini 3 ProGoogle28.0
50GLM-5Z.AI27.9
51Kimi K2.6Moonshot AI27.5
52MCQuasar 438BMultiverse Computing27.1
53Qwen3.6 PlusAlibaba27.0
54GLM-5-TurboZ.AI26.6
55GLM-5.1Z.AI26.4
56MiMo-V2.5-ProXiaomi26.4
57Claude Opus 4.6Anthropic26.4
58Kimi K2.7 CodeMoonshot AI26.3
59TMInkling-SmallThinking Machines Lab26.1
60Qwen3.7 PlusAlibaba25.8
61Hy3Tencent25.8
62TMInklingThinking Machines Lab25.5
63Ling 3.0 Flash VLInclusionAI25.0
64GPT-5.1OpenAI24.7
65Claude Sonnet 4.6Anthropic24.7
66GPT-5.4 miniOpenAI24.6
67MiMo-V2-OmniXiaomi23.9
68GPT-5.1-CodexOpenAI23.7
68GPT-5.1-Codex-MaxOpenAI23.7
70Claude Opus 4.5Anthropic23.7
71GLM-5V-TurboZ.AI23.5
72Kimi K2.5Moonshot AI23.5
72Kimi K2.5 (Reasoning)Moonshot AI23.5
74Nemotron 3 UltraNVIDIA23.4
75MiniMax M2.7MiniMax23.2
76GPT-5 (high)OpenAI23.0
77Qwen3.5-27BAlibaba22.9
78GPT-5 (medium)OpenAI22.9
79Claude 4.1 Opus ThinkingAnthropic22.9
80STA.X K2SK Telecom22.7
81Gemini 3.5 Flash-LiteGoogle22.7
82Command A+Cohere22.5
83Grok 4xAI22.5
84GLM-4.7Z.AI22.2
85Qwen3.6-27BAlibaba21.9
86o3-proOpenAI21.9
87Qwen3.5 397BAlibaba21.4
87Qwen3.5 397B (Reasoning)Alibaba21.4
89GPT-5.4 nanoOpenAI21.2
90Ling 3.0 FlashInclusionAI20.6
90Ling 3.0 Flash FP8InclusionAI20.6
92Grok 4.1 Fast (Reasoning)xAI20.4
93o3OpenAI20.2
94K-EXAONE 2.0LG AI Research19.7
95Step 3.7 FlashStepFun19.5
96Qwen3.5-35B-A3BAlibaba19.3
97Qwen3.6-35B-A3BAlibaba18.8
98Claude 4.1 OpusAnthropic18.6
99Muse Glimmer 30BMeta18.1
100Grok 4 Fast (Reasoning)xAI17.9
101Gemini 3 FlashGoogle17.9
102Gemma 4 26B A4BGoogle16.7
103Gemini 2.5 ProGoogle16.7
104Claude 4 SonnetAnthropic16.6
105Qwen3.5-122B-A10BAlibaba16.2
106MiMo-V2-FlashXiaomi16.1
107DeepSeek V3.2DeepSeek16.0
108Qwen3 MaxAlibaba15.6
109Gemma 4 31BGoogle15.4
110o1OpenAI15.2
111GLM-4.6Z.AI14.9
112Mistral Medium 3.5 128BMistral14.9
113Granite 4.2 30BIBM14.8
114K-ExaoneLG AI Research14.4
115Gemma 4 12BGoogle14.2
116Grok Code Fast 1xAI14.1
117Ling 2.6 FlashInclusionAI14.1
118DeepSeek V3.1DeepSeek13.7
119Nemotron 3.5 Lightning 30B A3B NVFP4NVIDIA13.6
120Nemotron 3 Super 100BNVIDIA13.6
121DeepSeek V3.1 (Reasoning)DeepSeek13.5
122OPMiniCPM5-2BOpenBMB13.1
123DeepSeek-R1DeepSeek13.1
124Kimi K2Moonshot AI12.7
125GPT-4.1OpenAI12.7
126o3-miniOpenAI12.5
127o1-proOpenAI12.4
128GPT-OSS 120BOpenAI12.3
129Ling 3.0 TinyInclusionAI11.9
130Granite 4.2 8BIBM11.8
131Mistral Small 4Mistral11.4
131Mistral Small 4 (Reasoning)Mistral11.4
133o1-previewOpenAI11.4
134Grok 4.1 FastxAI11.3
135GLM-4.5-AirZ.AI11.1
136Trinity-Large-PreviewArcee AI10.9
136Trinity-Large-ThinkingArcee AI10.9
138Nemotron 3 Nano Omni 30B A3BNVIDIA10.3
139GPT-4.1 miniOpenAI10.2
140North Mini CodeCohere9.9
141Gemini 2.5 FlashGoogle9.8
142Mistral Large 3Mistral9.7
143Llama 4 MaverickMeta9.3
144Granite 4.2 3BIBM9.1
145Mistral Medium 3Mistral9.1
146GPT-OSS 20BOpenAI9.0
147Gemma 4 E4BGoogle8.9
148Nemotron 3 Nano 30BNVIDIA8.9
149SASarvam 105BSarvam8.8
150Claude 3 OpusAnthropic8.7
151DeepSeek V3DeepSeek8.5
152GPT-4oOpenAI8.4
153LFM2.5-2.6BLiquidAI8.4
154DeepSeek R1 Distill Qwen 32BDeepSeek8.4
155Gemini 1.5 ProGoogle7.9
156GPT-4.1 nanoOpenAI7.8
156Solar Pro 3Upstage7.8
158Gemma 4 E2BGoogle7.8
159Qwen3-Omni-30B-A3B-ThinkingAlibaba7.8
160FAUltravox v0.6 Llama 3.3 70BFixie AI7.7
161Mistral Large 2Mistral7.6
162Nemotron Ultra 253BNVIDIA7.5
163Llama 3.1 405BMeta7.3
164LFM2.5-8B-A1BLiquidAI7.2
165GPT-4 TurboOpenAI7.0
166Solar Pro 2Upstage7.0
167Nova ProAmazon7.0
168Qwen2.5 Coder 32B InstructAlibaba6.7
169GPT-4o miniOpenAI6.7
170SASarvam 30BSarvam6.6
171Llama 4 ScoutMeta6.5
172CECeleris-1Celeris6.3
173Exaone 4.0 32BLG AI Research6.3
174Qwen3-Omni-30B-A3B-InstructAlibaba6.0
175Phi-4Microsoft5.9
176Phi-4 Multimodal InstructMicrosoft5.8
177Claude 3 HaikuAnthropic5.6
178Gemini 1.0 ProGoogle5.3
179Exaone 4.0 1.2BLG AI Research5.2
180Granite-4.0-H-1BIBM5.2
181Granite-4.0-1BIBM5.0
182Gemma 3 27BGoogle4.8
183Granite-4.0-350MIBM4.8
183Granite-4.0-H-350MIBM4.8
183LFM2.5-VL-1.6B-ExtractLiquidAI4.8

Evidence key: ObservedLast good

Rows are ordered by the value Artificial Analysis published, highest first. The source does not state whether a higher value is the better result, so this page does not either: for a benchmark that measures a rate of failure, read the table from the bottom.

All leaderboards · Knowledge capability leaderboard