Skip to main content
BenchLM
Data

Artificial Analysis Intelligence Index

We show this table for reference; we do not rank on it.

GPT-5.6 Sol has the highest reported Artificial Analysis Intelligence Index score at 58.9%, ahead of Claude Opus 5.5 (57.6%) and Claude Sonnet 5.5 (56.0%) among 218 models. We compile the table from provider self-reports and secondary reports and keep it for display only; it does not affect overall rankings.

Data verified 38 confirmed releases in the last 30 daysFollow model changes

A display-only intelligence index published by Artificial Analysis that aggregates provider-reported and benchmark-derived signals into a single model-level score.

Benchmark score on Artificial Analysis Intelligence Index — October 7, 2026

We compile the Artificial Analysis Intelligence Index rows from provider self-reports and secondary reports. GPT-5.6 Sol leads the table at 58.9%, followed by Claude Opus 5.5 (57.6%) and Claude Sonnet 5.5 (56.0%). We do not use these results to rank models overall.

218 modelsKnowledgeCurrentDisplay onlyUpdated October 7, 2026

Benchmark score results for Artificial Analysis Intelligence Index
RankModel / configurationScoreParameters (B)Open / closed
1GPT-5.6 SolOpenAI
58.9%
Not reportedClosed
2Claude Opus 5.5Anthropic
57.6%
Not reportedClosed
3Claude Sonnet 5.5Anthropic
56.0%
Not reportedClosed
4GPT-5.6 TerraOpenAI
55.0%
Not reportedClosed
5Claude Fable 5.1Anthropic
53.4%
Not reportedClosed
6DeepSeek V4 Pro 0813DeepSeek
53.2%
Not reportedOpen
7GPT-6 AstraOpenAI
52.7%
Not reportedClosed
8Gemini 4 ArgonGoogle
52.6%
Not reportedClosed
9GPT-6.1 SolOpenAI
51.8%
Not reportedClosed
10GPT-5.6 LunaOpenAI
51.2%
Not reportedClosed
11Claude Opus 5Anthropic
50.8%
Not reportedClosed
12Gemini 3.5 FlashGoogle
50.2%
Not reportedClosed
13Claude Fable 5Anthropic
49.6%
Not reportedClosed
14Muse Spark 1.3Meta
48.1%
Not reportedClosed
15GPT-6 SolOpenAI
47.6%
Not reportedClosed
16Grok 4.7xAI
46.5%
Not reportedClosed
17MiMo-V2.6-ProXiaomi
46.3%
Not reportedOpen
18Qwen3.8 Max PreviewAlibaba
45.4%
Not reportedClosed
19GLM-5.3Z.AI
44.8%
Not reportedOpen
20Grok 4.6xAI
44.3%
Not reportedClosed
21DeepSeek V4 Pro (High)DeepSeek
43.8%
Not reportedOpen
22Step 5 PreviewStepFun
43.7%
Not reportedPending
23Kimi K3Moonshot AI
43.6%
Not reportedPending
24Claude Haiku 5.5Anthropic
43.4%
Not reportedClosed
25GLM-5.3-FlashZ.AI
41.8%
Not reportedOpen
26Claude Opus 4.8Anthropic
41.8%
Not reportedClosed
27Hy3 PreviewTencent
41.2%
Not reportedOpen
28Ling 3.1 FlashInclusionAI
41.1%
Not reportedClosed
29Gemini 3.8 FlashGoogle
40.9%
Not reportedClosed
30Claude Opus 4.7 (Adaptive)Anthropic
40.7%
Not reportedClosed
31Qwen3.8-Flash-NextAlibaba
39.8%
Not reportedOpen
32Muse Spark 1.2Meta
39.6%
Not reportedClosed
33DeepSeek V4.1 FlashDeepSeek
39.5%
Not reportedOpen
34Gemini 3.7 FlashGoogle
39.1%
Not reportedClosed
35GPT-5.4OpenAI
39.0%
Not reportedClosed
36Grok 4.5xAI
38.8%
Not reportedClosed
37Mistral Large 4Mistral
38.4%
Not reportedPending
38GPT-5.5OpenAI
38.4%
Not reportedClosed
39Claude Sonnet 5Anthropic
38.2%
Not reportedClosed
40GPT-6 LunaOpenAI
38.1%
Not reportedClosed
41MiMo-V2.6-FlashXiaomi
37.9%
Not reportedOpen
42Grok 4.3xAI
37.6%
Not reportedClosed
43DeepSeek V4 Flash 0731DeepSeek
34.3%
Not reportedOpen
44Gemini 3.6 FlashGoogle
34.0%
Not reportedClosed
45Muse Spark 1.1Meta
33.7%
Not reportedClosed
46GLM-5.2Z.AI
33.7%
Not reportedOpen
47Qwen3.8-27BAlibaba
33.7%
Not reportedOpen
48GPT-5.3 CodexOpenAI
32.5%
Not reportedClosed
49GPT-5.3-Codex-SparkOpenAI
32.5%
Not reportedClosed
50Claude Opus 4.6 (Adaptive)Anthropic
31.9%
Not reportedClosed
51Muse SparkMeta
31.3%
Not reportedClosed
52Claude Opus 4.7Anthropic
30.9%
Not reportedClosed
53GPT-5.2OpenAI
30.4%
Not reportedClosed
54Gemini 3.1 ProGoogle
29.7%
Not reportedClosed
55Qwen3.7 MaxAlibaba
29.5%
Not reportedClosed
56MiniMax M3MiniMax
29.2%
Not reportedOpen
57Claude Opus 4.5 ThinkingAnthropic
29.1%
Not reportedClosed
58MiMo-V2-ProXiaomi
28.6%
Not reportedClosed
59GPT-5.2-CodexOpenAI
28.5%
Not reportedClosed
60Qwen 3.6 Max (preview)Alibaba
28.4%
Not reportedClosed
61Solar Pro 4Upstage
28.1%
Not reportedClosed
62Gemini 3 ProGoogle
28.0%
Not reportedClosed
63GLM-5Z.AI
27.9%
Not reportedOpen
64Qwen3.6 PlusAlibaba
27.0%
Not reportedClosed
65Kimi K2.6Moonshot AI
27.0%
Not reportedOpen
66Quasar 438BMultiverse Computing
26.7%
Not reportedClosed
67GLM-5-TurboZ.AI
26.6%
Not reportedClosed
68Apodex 1.1Apodex
26.4%
Not reportedClosed
69Apodex 1.1 MiniApodex
26.4%
Not reportedOpen
70Claude Opus 4.6Anthropic
26.4%
Not reportedClosed
71GLM-5.1Z.AI
26.1%
Not reportedOpen
72MiMo-V2.5-ProXiaomi
26.0%
Not reportedClosed
73Kimi K2.7 CodeMoonshot AI
25.8%
Not reportedOpen
74Inkling-SmallThinking Machines Lab
25.7%
Not reportedOpen
75Hy3Tencent
25.3%
Not reportedOpen
76Qwen3.7 PlusAlibaba
25.2%
Not reportedClosed
77InklingThinking Machines Lab
25.0%
Not reportedOpen
78GPT-5.1OpenAI
24.7%
Not reportedClosed
79Claude Sonnet 4.6Anthropic
24.7%
Not reportedClosed
80Ling 3.0 Flash VLInclusionAI
24.6%
Not reportedOpen
81GPT-5.4 miniOpenAI
24.1%
Not reportedClosed
82MiMo-V2-OmniXiaomi
23.9%
Not reportedClosed
83GPT-5.1-CodexOpenAI
23.7%
Not reportedClosed
84GPT-5.1-Codex-MaxOpenAI
23.7%
Not reportedClosed
85Claude Opus 4.5Anthropic
23.7%
Not reportedClosed
86GLM-5V-TurboZ.AI
23.5%
Not reportedClosed
87Kimi K2.5Moonshot AI
23.5%
Not reportedOpen
88Kimi K2.5 (Reasoning)Moonshot AI
23.5%
Not reportedClosed
89GPT-5 (high)OpenAI
23.0%
Not reportedClosed
90Nemotron 3 UltraNVIDIA
22.9%
Not reportedOpen
91Qwen3.5-27BAlibaba
22.9%
Not reportedOpen
92GPT-5 (medium)OpenAI
22.9%
Not reportedClosed
93Claude 4.1 Opus ThinkingAnthropic
22.9%
Not reportedClosed
94MiniMax M2.5MiniMax
22.8%
Not reportedClosed
95MiniMax M2.7MiniMax
22.8%
Not reportedOpen
96Command A+Cohere
22.5%
Not reportedOpen
97Grok 4xAI
22.5%
Not reportedClosed
98GLM-4.7Z.AI
22.2%
Not reportedOpen
99Gemini 3.5 Flash-LiteGoogle
22.2%
Not reportedClosed
100o3-proOpenAI
21.9%
Not reportedClosed
101A.X K2SK Telecom
21.5%
Not reportedOpen
102Qwen3.5 397BAlibaba
21.4%
Not reportedOpen
103Qwen3.5 397B (Reasoning)Alibaba
21.4%
Not reportedOpen
104Qwen3.6-27BAlibaba
21.4%
Not reportedOpen
105GPT-5.4 nanoOpenAI
20.7%
Not reportedClosed
106Grok 4.1 Fast (Reasoning)xAI
20.4%
Not reportedClosed
107o3OpenAI
20.2%
Not reportedClosed
108Ling 3.0 FlashInclusionAI
20.1%
Not reportedOpen
109Ling 3.0 Flash FP8InclusionAI
20.1%
Not reportedOpen
110K-EXAONE 2.0LG AI Research
19.7%
Not reportedOpen
111Step 3.7 FlashStepFun
19.5%
Not reportedOpen
112Qwen3.5-35B-A3BAlibaba
19.3%
Not reportedOpen
113Claude 4.1 OpusAnthropic
18.6%
Not reportedClosed
114Qwen3.6-35B-A3BAlibaba
18.2%
Not reportedOpen
115Grok 4 Fast (Reasoning)xAI
17.9%
Not reportedClosed
116Gemini 3 FlashGoogle
17.9%
Not reportedClosed
117Muse Glimmer 30BMeta
17.5%
Not reportedOpen
118Step 3.5 FlashStepFun
17.0%
Not reportedOpen
119GPT-5 miniOpenAI
16.8%
Not reportedClosed
120Gemma 4 26B A4BGoogle
16.7%
Not reportedOpen
121Claude 4 SonnetAnthropic
16.6%
Not reportedClosed
122Gemini 2.5 ProGoogle
16.1%
Not reportedClosed
123MiMo-V2-FlashXiaomi
16.1%
Not reportedOpen
124DeepSeek V3.2DeepSeek
16.0%
Not reportedOpen
125Qwen3 MaxAlibaba
15.6%
Not reportedClosed
126Qwen3.5-122B-A10BAlibaba
15.6%
Not reportedOpen
127o1OpenAI
15.2%
Not reportedClosed
128GLM-4.6Z.AI
14.9%
Not reportedOpen
129GLM-4.7-FlashZ.AI
14.9%
Not reportedOpen
130Granite 4.2 30BIBM
14.8%
Not reportedOpen
131Gemma 4 31BGoogle
14.7%
Not reportedOpen
132K-ExaoneLG AI Research
14.4%
Not reportedClosed
133Mistral Medium 3.5 128BMistral
14.2%
Not reportedOpen
134Gemma 4 12BGoogle
14.2%
Not reportedOpen
135Grok Code Fast 1xAI
14.1%
Not reportedClosed
136Ling 2.6 FlashInclusionAI
14.1%
Not reportedOpen
137Mercury 2Inception
13.8%
Not reportedClosed
138DeepSeek V3.1DeepSeek
13.7%
Not reportedOpen
139DeepSeek V3.1 (Reasoning)DeepSeek
13.5%
Not reportedOpen
140DeepSeek-R1DeepSeek
13.1%
Not reportedOpen
141GPT-5 nanoOpenAI
13.0%
Not reportedClosed
142Nemotron 3.5 Lightning 30B A3B NVFP4NVIDIA
12.9%
Not reportedOpen
143Nemotron 3 Super 100BNVIDIA
12.8%
Not reportedOpen
144Nemotron 3 Super 120B A12BNVIDIA
12.8%
Not reportedOpen
145Kimi K2Moonshot AI
12.7%
Not reportedClosed
146GPT-4.1OpenAI
12.7%
Not reportedClosed
147o3-miniOpenAI
12.5%
Not reportedClosed
148MiniCPM5-2BOpenBMB
12.5%
Not reportedOpen
149o1-proOpenAI
12.4%
Not reportedClosed
150Mercury 2.5Inception
12.3%
Not reportedClosed
151MiniMax M1 80kMiniMax
11.7%
Not reportedClosed
152GPT-OSS 120BOpenAI
11.6%
Not reportedOpen
153o1-previewOpenAI
11.4%
Not reportedClosed
154Grok 4.1 FastxAI
11.3%
Not reportedClosed
155Mistral Small 4Mistral
11.3%
Not reportedOpen
156Mistral Small 4 (Reasoning)Mistral
11.3%
Not reportedOpen
157GLM-4.5-AirZ.AI
11.1%
Not reportedClosed
158Granite 4.2 8BIBM
11.1%
Not reportedOpen
159Ling 3.0 TinyInclusionAI
11.1%
Not reportedOpen
160Trinity-Large-ThinkingArcee AI
10.8%
Not reportedOpen
161Trinity-Large-PreviewArcee AI
10.8%
Not reportedOpen
162Nemotron 3 Nano Omni 30B A3BNVIDIA
10.3%
Not reportedOpen
163GPT-4.1 miniOpenAI
10.2%
Not reportedClosed
164Llama 4 MaverickMeta
10.0%
Not reportedOpen
165North Mini CodeCohere
9.9%
Not reportedOpen
166Gemini 2.5 FlashGoogle
9.8%
Not reportedClosed
167DeepSeek V3 0324DeepSeek
9.7%
Not reportedOpen
168Mistral Large 3Mistral
9.3%
Not reportedClosed
169Granite 4.2 3BIBM
9.1%
Not reportedOpen
170Mistral Medium 3Mistral
9.1%
Not reportedClosed
171GPT-OSS 20BOpenAI
9.0%
Not reportedOpen
172Gemma 4 E4BGoogle
8.9%
Not reportedOpen
173Nemotron 3 Nano 30BNVIDIA
8.9%
Not reportedOpen
174Sarvam 105BSarvam
8.8%
Not reportedOpen
175Claude 3 OpusAnthropic
8.7%
Not reportedClosed
176DeepSeek V3DeepSeek
8.5%
Not reportedOpen
177GPT-4oOpenAI
8.4%
Not reportedClosed
178LFM2.5-2.6BLiquidAI
8.4%
Not reportedOpen
179DeepSeek R1 Distill Qwen 32BDeepSeek
8.4%
Not reportedOpen
180Llama 4 ScoutMeta
8.1%
Not reportedOpen
181Gemini 1.5 ProGoogle
7.9%
Not reportedClosed
182Solar Pro 3Upstage
7.8%
Not reportedClosed
183GPT-4.1 nanoOpenAI
7.8%
Not reportedClosed
184Gemma 4 E2BGoogle
7.8%
Not reportedOpen
185Qwen3-Omni-30B-A3B-ThinkingAlibaba
7.8%
Not reportedOpen
186Ultravox v0.6 Llama 3.3 70BFixie AI
7.7%
Not reportedOpen
187Mistral Large 2Mistral
7.6%
Not reportedClosed
188Nemotron Ultra 253BNVIDIA
7.5%
Not reportedOpen
189Llama 3.1 405BMeta
7.3%
Not reportedOpen
190LFM2.5-8B-A1BLiquidAI
7.2%
Not reportedOpen
191GPT-4 TurboOpenAI
7.0%
Not reportedClosed
192Solar Pro 2Upstage
7.0%
Not reportedClosed
193Nova ProAmazon
7.0%
Not reportedClosed
194Qwen2.5 Coder 32B InstructAlibaba
6.7%
Not reportedOpen
195GPT-4o miniOpenAI
6.7%
Not reportedClosed
196Sarvam 30BSarvam
6.6%
Not reportedOpen
197Celeris-1Celeris
6.3%
Not reportedClosed
198Exaone 4.0 32BLG AI Research
6.3%
Not reportedOpen
199Ministral 3 14B (Reasoning)Mistral
6.0%
Not reportedOpen
200Ministral 3 14BMistral
6.0%
Not reportedOpen
201Qwen3-Omni-30B-A3B-InstructAlibaba
6.0%
Not reportedOpen
202LFM2-24B-A2BLiquidAI
6.0%
Not reportedClosed
203Phi-4Microsoft
5.9%
Not reportedOpen
204Phi-4 Multimodal InstructMicrosoft
5.8%
Not reportedOpen
205Claude 3 HaikuAnthropic
5.6%
Not reportedClosed
206Ministral 3 8B (Reasoning)Mistral
5.5%
Not reportedOpen
207Ministral 3 8BMistral
5.5%
Not reportedOpen
208Gemini 1.0 ProGoogle
5.3%
Not reportedClosed
209Exaone 4.0 1.2BLG AI Research
5.2%
Not reportedOpen
210LFM2.5-1.2B-ThinkingLiquidAI
5.2%
Not reportedClosed
211LFM2.5-1.2B-InstructLiquidAI
5.2%
Not reportedClosed
212Granite-4.0-H-1BIBM
5.2%
Not reportedOpen
213Gemma 3 27BGoogle
4.8%
Not reportedOpen
214Ministral 3 3B (Reasoning)Mistral
4.8%
Not reportedOpen
215Ministral 3 3BMistral
4.8%
Not reportedOpen
216Granite-4.0-350MIBM
4.8%
Not reportedOpen
217Granite-4.0-H-350MIBM
4.8%
Not reportedOpen
218LFM2.5-VL-1.6B-ExtractLiquidAI
4.8%
Not reportedOpen

Among the reported Artificial Analysis Intelligence Index rows, GPT-5.6 Sol is first at 58.9%. The third row is 2.9 points behind. The broader top-10 range is 7.6 points, so many of the published results sit in a relatively narrow band.

218 models have been evaluated on Artificial Analysis Intelligence Index. The benchmark falls in the Knowledge category. Artificial Analysis Intelligence Index is currently displayed for reference but excluded from the scoring formula, so it does not directly affect overall rankings.

About Artificial Analysis Intelligence Index

Year

2026

Tasks

Cross-benchmark intelligence index

Format

Aggregated model score

Difficulty

Display-only external reference

BenchLM tracks Artificial Analysis as a display-only external reference rather than a weighted benchmark. It is useful as a market snapshot, but it is not a benchmark-native row with a single public task set, scoring harness, or exact-source methodology aligned to BenchLM's core benchmark pages.

Freshness and provenance

Version

Artificial Analysis Intelligence Index 2026

Refresh cadence

Quarterly

Staleness state

Current

Question availability

Public benchmark set

CurrentDisplay only

BenchLM uses freshness metadata to decide whether a benchmark should still be treated as a strong differentiator, a benchmark to watch, or a display-only reference. For the full scoring policy, see the BenchLM methodology page.

Questions

What does Artificial Analysis Intelligence Index measure?

A display-only intelligence index published by Artificial Analysis that aggregates provider-reported and benchmark-derived signals into a single model-level score.

Which model scores highest on Artificial Analysis Intelligence Index?

GPT-5.6 Sol by OpenAI currently leads with a score of 58.9% on Artificial Analysis Intelligence Index.

How many models are evaluated on Artificial Analysis Intelligence Index?

218 AI models have published results on Artificial Analysis Intelligence Index in the BenchLM catalog.

Last updated: October 7, 2026 · BenchLM version Artificial Analysis Intelligence Index 2026

Know when it’s worth switching models

The model to choose, the cheaper alternative, and the release we would wait on.

Read a sample issue

Join 5,500+ readers.

One email each week. Unsubscribe anytime.