Skip to main content
Radar

Every change to the models you run, with its source and its date. Releases, price changes, retirements, API changes, and incidents.Every change to the models you run, with its source.

Follow model changes

Artificial Analysis Omniscience Index (AA-Omniscience Index)

A display-only Artificial Analysis factual knowledge index.

Data verified 33 confirmed releases in the last 30 daysSee provider release alerts

Benchmark score on AA-Omniscience Index — September 15, 2026

We mirror the published score view for AA-Omniscience Index. Claude Fable 5.1 leads the public snapshot at 43.5%, followed by GPT-6 Astra (43.4%) and Claude Fable 5 (43.3%). We do not use these results to rank models overall.

190 modelsKnowledgeCurrentDisplay onlyUpdated September 15, 2026

Benchmark score table (190 models)

Score
1
Claude Fable 5.1Anthropic · Closed
43.5%
2
GPT-6 AstraOpenAI · Closed
43.4%
3
Claude Fable 5Anthropic · Closed
43.3%
4
Claude Opus 5Anthropic · Closed
37.1%
5
Gemini 3.1 ProGoogle · Closed
31.9%
6
Grok 4.6xAI · Closed
30.5%
7
Gemini 3.8 FlashGoogle · Closed
29.6%
8
Claude Opus 4.8Anthropic · Closed
28.8%
9
Muse Spark 1.1Meta · Closed
28.1%
10
Claude Opus 4.7 (Adaptive)Anthropic · Closed
27.3%
11
Muse Spark 1.2Meta · Closed
27.2%
12
Gemini 3.7 FlashGoogle · Closed
26.5%
13
Grok 4.5xAI · Closed
25.3%
14
Muse Spark 1.3Meta · Closed
25.0%
15
Gemini 3.6 FlashGoogle · Closed
22.1%
16
GPT-5.6 SolOpenAI · Closed
22.0%
17
Gemini 3.5 FlashGoogle · Closed
21.2%
18
GPT-5.5OpenAI · Closed
20.5%
19
Kimi K3Moonshot AI · Closed
19.7%
20
Grok 4.3xAI · Closed
18.0%
21
Claude Sonnet 5Anthropic · Closed
16.5%
22
Gemini 3 ProGoogle · Closed
15.3%
23
Claude Opus 4.7Anthropic · Closed
14.8%
24
GLM-5.3Z.AI · Open weight
14.3%
25
Claude Opus 4.5 ThinkingAnthropic · Closed
14.0%
26
Claude Opus 4.6 (Adaptive)Anthropic · Closed
13.7%
27
Qwen3.7 MaxAlibaba · Closed
13.5%
28
GPT-5.3 CodexOpenAI · Closed
10.9%
29
GPT-5.3-Codex-SparkOpenAI · Closed
10.9%
30
Qwen 3.6 Max (preview)Alibaba · Closed
9.2%
31
GLM-5.3-FlashZ.AI · Open weight
7.5%
32
Muse SparkMeta · Closed
7.2%
33
GPT-5.4OpenAI · Closed
5.8%
34
GPT-5.1OpenAI · Closed
5.4%
35
Kimi K2.6Moonshot AI · Open weight
5.3%
36
Gemini 3.5 Flash-LiteGoogle · Closed
5.2%
37
MiMo-V2-ProXiaomi · Closed
4.6%
38
GLM-5.2Z.AI · Open weight
4.4%
39
Qwen3.8 Max PreviewAlibaba · Closed
3.4%
40
MiMo-V2.5-ProXiaomi · Closed
3.3%
41
Claude Opus 4.6Anthropic · Closed
2.4%
42
Grok 4xAI · Closed
2.1%
43
InklingThinking Machines Lab · Open weight
2.0%
44
MiniMax M3MiniMax · Open weight
1.4%
45
Qwen3.7 PlusAlibaba · Closed
1.1%
46
GLM-5.1Z.AI · Open weight
0.9%
47
Qwen3.6 PlusAlibaba · Closed
0.9%
48
DeepSeek V4 Pro 0813DeepSeek · Closed
0.8%
49
MiniMax M2.7MiniMax · Open weight
0.8%
50
GLM-5Z.AI · Open weight
0.3%
51
GPT-5.6 TerraOpenAI · Closed
0.1%
52
Nemotron 3 UltraNVIDIA · Open weight
-0.4%
53
Solar Pro 4Upstage · Closed
-0.8%
54
GPT-5.2OpenAI · Closed
-0.9%
55
GPT-5.2-CodexOpenAI · Closed
-2.2%
56
Quasar 438BMultiverse Computing · Closed
-2.6%
57
Claude Sonnet 4.6Anthropic · Closed
-3.5%
58
Command A+Cohere · Open weight
-4.0%
59
Claude Opus 4.5Anthropic · Closed
-4.1%
60
Gemini 3 FlashGoogle · Closed
-4.3%
61
Ling 3.0 Flash VLInclusionAI · Open weight
-4.5%
62
DeepSeek V4.1 FlashDeepSeek · Open weight
-5.3%
63
GPT-5.1-CodexOpenAI · Closed
-6.5%
64
GPT-5.1-Codex-MaxOpenAI · Closed
-6.5%
65
K-EXAONE 2.0LG AI Research · Open weight
-6.6%
66
Kimi K2.5Moonshot AI · Open weight
-7.3%
67
Kimi K2.5 (Reasoning)Moonshot AI · Closed
-7.3%
68
A.X K2SK Telecom · Open weight
-8.2%
69
GPT-5 (high)OpenAI · Closed
-8.7%
70
Inkling-SmallThinking Machines Lab · Open weight
-8.9%
71
Claude 4 SonnetAnthropic · Closed
-9.0%
72
Qwen3.8-Flash-NextAlibaba · Open weight
-9.7%
73
Qwen3.8-27BAlibaba · Open weight
-10.0%
74
Kimi K2.7 CodeMoonshot AI · Open weight
-10.2%
75
GPT-5.6 LunaOpenAI · Closed
-10.3%
76
GPT-4oOpenAI · Closed
-10.5%
77
DeepSeek V4 Pro (High)DeepSeek · Open weight
-10.6%
78
GPT-5 (medium)OpenAI · Closed
-10.9%
79
o1OpenAI · Closed
-11.0%
80
MiniCPM5-2BOpenBMB · Open weight
-11.6%
81
Granite 4.2 30BIBM · Open weight
-12.9%
82
DeepSeek V4 Flash 0731DeepSeek · Closed
-14.3%
83
Granite 4.2 3BIBM · Open weight
-14.7%
84
o3OpenAI · Closed
-15.5%
85
Gemini 2.5 ProGoogle · Closed
-16.3%
86
GLM-5-TurboZ.AI · Closed
-16.4%
87
Llama 3.1 405BMeta · Open weight
-17.1%
88
Granite 4.2 8BIBM · Open weight
-17.2%
89
GPT-5 miniOpenAI · Closed
-17.3%
90
-17.7%
91
Ling 3.0 FlashInclusionAI · Open weight
-17.9%
92
Ling 3.0 Flash FP8InclusionAI · Open weight
-17.9%
93
Hy3Tencent · Open weight
-18.5%
94
Hy3 PreviewTencent · Open weight
-18.5%
95
GPT-5.4 miniOpenAI · Closed
-18.9%
96
GLM-5V-TurboZ.AI · Closed
-19.3%
97
Ling 3.0 TinyInclusionAI · Open weight
-19.3%
98
Gemma 4 E4BGoogle · Open weight
-19.7%
99
Qwen3.6-27BAlibaba · Open weight
-20.0%
100
MiMo-V2-OmniXiaomi · Closed
-20.1%
101
Apodex 1.1Apodex · Closed
-21.9%
102
Apodex 1.1 MiniApodex · Open weight
-21.9%
103
Qwen3.6-35B-A3BAlibaba · Open weight
-22.2%
104
Gemma 4 E2BGoogle · Open weight
-23.6%
105
DeepSeek-R1DeepSeek · Open weight
-27.4%
106
Kimi K2Moonshot AI · Closed
-28.3%
107
GPT-5 nanoOpenAI · Closed
-28.7%
108
GPT-5.4 nanoOpenAI · Closed
-29.5%
109
LFM2.5-2.6BLiquidAI · Open weight
-29.5%
110
DeepSeek V3.1 (Reasoning)DeepSeek · Open weight
-29.6%
111
-29.9%
112
-29.9%
113
Mistral Small 4Mistral · Open weight
-30.4%
114
Mistral Small 4 (Reasoning)Mistral · Open weight
-30.4%
115
Mistral Medium 3Mistral · Closed
-31.4%
116
GLM-4.6Z.AI · Open weight
-31.7%
117
Muse Glimmer 30BMeta · Open weight
-32.8%
118
LFM2.5-8B-A1BLiquidAI · Open weight
-33.3%
119
Mistral Large 2Mistral · Closed
-34.4%
120
GLM-4.7Z.AI · Open weight
-36.4%
121
Mistral Medium 3.5 128BMistral · Open weight
-36.8%
122
Grok Code Fast 1xAI · Closed
-37.1%
123
Step 3.7 FlashStepFun · Open weight
-37.3%
124
Qwen3.5 397BAlibaba · Open weight
-37.9%
125
Qwen3.5 397B (Reasoning)Alibaba · Open weight
-37.9%
126
MiniMax M2.5MiniMax · Closed
-38.9%
127
Mistral Large 3Mistral · Closed
-39.6%
128
GPT-4.1OpenAI · Closed
-39.6%
129
Qwen3.5-122B-A10BAlibaba · Open weight
-41.5%
130
Nemotron 3 Super 100BNVIDIA · Open weight
-41.5%
131
Nemotron 3 Super 120B A12BNVIDIA · Open weight
-41.5%
132
DeepSeek V3DeepSeek · Open weight
-41.6%
133
Llama 4 MaverickMeta · Open weight
-41.8%
134
Gemini 2.5 FlashGoogle · Closed
-42.6%
135
DeepSeek V3.1DeepSeek · Open weight
-42.7%
136
Qwen3 MaxAlibaba · Closed
-43.5%
137
Qwen3.5-27BAlibaba · Open weight
-44.0%
138
Trinity-Large-ThinkingArcee AI · Open weight
-44.1%
139
Trinity-Large-PreviewArcee AI · Open weight
-44.1%
140
Step 3.5 FlashStepFun · Open weight
-44.2%
141
Nemotron Ultra 253BNVIDIA · Open weight
-44.9%
142
DeepSeek V3.2DeepSeek · Open weight
-46.9%
143
Nova ProAmazon · Closed
-47.7%
144
Gemma 4 31BGoogle · Open weight
-47.9%
145
Qwen3.5-35B-A3BAlibaba · Open weight
-48.1%
146
MiMo-V2-FlashXiaomi · Open weight
-48.4%
147
MiniMax M1 80kMiniMax · Closed
-48.5%
148
Claude 3 HaikuAnthropic · Closed
-48.6%
149
North Mini CodeCohere · Open weight
-48.6%
150
GPT-OSS 120BOpenAI · Open weight
-49.2%
151
Mercury 2Inception · Closed
-50.7%
152
Gemma 4 26B A4BGoogle · Open weight
-50.8%
153
Grok 4.1 FastxAI · Closed
-50.9%
154
Nemotron 3 Nano 30BNVIDIA · Open weight
-51.6%
155
Llama 4 ScoutMeta · Open weight
-52.1%
156
Gemma 4 12BGoogle · Open weight
-52.7%
157
Solar Pro 3Upstage · Closed
-53.4%
158
GPT-4.1 miniOpenAI · Closed
-53.6%
159
Ultravox v0.6 Llama 3.3 70BFixie AI · Open weight
-54.2%
160
Phi-4Microsoft · Open weight
-55.7%
161
Nemotron 3 Nano Omni 30B A3BNVIDIA · Open weight
-57.4%
162
GPT-4.1 nanoOpenAI · Closed
-57.6%
163
K-ExaoneLG AI Research · Closed
-58.0%
164
LFM2-24B-A2BLiquidAI · Closed
-58.1%
165
Sarvam 105BSarvam · Open weight
-59.4%
166
Qwen3-Omni-30B-A3B-ThinkingAlibaba · Open weight
-61.4%
167
GLM-4.5-AirZ.AI · Closed
-61.5%
168
Solar Pro 2Upstage · Closed
-62.0%
169
GLM-4.7-FlashZ.AI · Open weight
-62.6%
170
Exaone 4.0 32BLG AI Research · Open weight
-62.8%
171
GPT-OSS 20BOpenAI · Open weight
-63.0%
172
Ministral 3 3B (Reasoning)Mistral · Open weight
-64.0%
173
Ministral 3 3BMistral · Open weight
-64.0%
174
Ling 2.6 FlashInclusionAI · Open weight
-66.1%
175
Ministral 3 14B (Reasoning)Mistral · Open weight
-66.4%
176
Ministral 3 14BMistral · Open weight
-66.4%
177
Gemma 3 27BGoogle · Open weight
-67.2%
178
Ministral 3 8B (Reasoning)Mistral · Open weight
-68.9%
179
Ministral 3 8BMistral · Open weight
-68.9%
180
Qwen3-Omni-30B-A3B-InstructAlibaba · Open weight
-69.3%
181
Granite-4.0-350MIBM · Open weight
-69.3%
182
Sarvam 30BSarvam · Open weight
-71.5%
183
Celeris-1Celeris · Closed
-71.6%
184
LFM2.5-1.2B-InstructLiquidAI · Closed
-72.1%
185
Granite-4.0-H-1BIBM · Open weight
-72.3%
186
LFM2.5-1.2B-ThinkingLiquidAI · Closed
-79.9%
187
Granite-4.0-H-350MIBM · Open weight
-80.9%
188
Granite-4.0-1BIBM · Open weight
-81.6%
189
Exaone 4.0 1.2BLG AI Research · Open weight
-82.1%
190
LFM2.5-VL-1.6B-ExtractLiquidAI · Open weight
-84.4%

The published AA-Omniscience Index snapshot places Claude Fable 5.1 first at 43.5%. The third row is 0.2 points behind. The broader top-10 range is 16.2 points, so the table still separates the published systems.

190 models have been evaluated on AA-Omniscience Index. The benchmark falls in the Knowledge category. This category carries a 12% weight in BenchLM.ai's overall scoring system. AA-Omniscience Index is currently displayed for reference but excluded from the scoring formula, so it does not directly affect overall rankings.

About AA-Omniscience Index

Year

2026

Tasks

Knowledge questions

Format

Index score

Difficulty

Broad factual knowledge

BenchLM stores the AA-Omniscience index as a display-only factuality signal alongside the accuracy and hallucination-rate rows.

BenchLM freshness & provenance

Version

AA-Omniscience Index 2026

Refresh cadence

Quarterly

Staleness state

Current

Question availability

Public benchmark set

CurrentDisplay only

BenchLM uses freshness metadata to decide whether a benchmark should still be treated as a strong differentiator, a benchmark to watch, or a display-only reference. For the full scoring policy, see the BenchLM methodology page.

FAQ

What does AA-Omniscience Index measure?

A display-only Artificial Analysis factual knowledge index.

Which model scores highest on AA-Omniscience Index?

Claude Fable 5.1 by Anthropic currently leads with a score of 43.5% on AA-Omniscience Index.

How many models are evaluated on AA-Omniscience Index?

190 AI models have been evaluated on AA-Omniscience Index on BenchLM.

Last updated: September 15, 2026 · BenchLM version AA-Omniscience Index 2026

Know when it’s worth switching models

The model to choose, the cheaper alternative, and the release we would wait on.

Read a sample issue

Join 2,000+ readers.

One email each week. Unsubscribe anytime.