Artificial Analysis Omniscience Hallucination Rate (AA-Omniscience Hallucination Rate)
We mirror this table; we do not rank on it.
A display-only Artificial Analysis factuality metric for the rate of incorrect answers among non-correct responses.
Benchmark score on AA-Omniscience Hallucination Rate — September 18, 2026
We mirror the published score view for AA-Omniscience Hallucination Rate. Command A+ leads the public snapshot at 14.2%, followed by LFM2.5-2.6B (16.0%) and MiniMax M3 (18.4%). We do not use these results to rank models overall.
Command A+
Cohere
LFM2.5-2.6B
LiquidAI
MiniMax M3
MiniMax
189 modelsKnowledgeCurrentDisplay onlyUpdated September 18, 2026
Benchmark score table (189 models)
ScoreThe published AA-Omniscience Hallucination Rate snapshot places Command A+ first at 14.2%. The third row is 4.2 points higher. The broader top-10 range is 10.8 points, so the table still separates the published systems.
189 models have been evaluated on AA-Omniscience Hallucination Rate. The benchmark falls in the Knowledge category. This category carries a 12% weight in BenchLM.ai's overall scoring system. AA-Omniscience Hallucination Rate is currently displayed for reference but excluded from the scoring formula, so it does not directly affect overall rankings.
About AA-Omniscience Hallucination Rate
Year
2026
Tasks
Knowledge questions
Format
Hallucination rate
Difficulty
Factuality
BenchLM marks this row lower-is-better because a lower hallucination rate is preferable, even though the OpenRouter card displays the raw percentage.
Freshness and provenance
Version
AA-Omniscience Hallucination Rate 2026
Refresh cadence
Quarterly
Staleness state
Current
Question availability
Public benchmark set
BenchLM uses freshness metadata to decide whether a benchmark should still be treated as a strong differentiator, a benchmark to watch, or a display-only reference. For the full scoring policy, see the BenchLM methodology page.
Questions
What does AA-Omniscience Hallucination Rate measure?
A display-only Artificial Analysis factuality metric for the rate of incorrect answers among non-correct responses.
Which model scores highest on AA-Omniscience Hallucination Rate?
Command A+ by Cohere currently leads with a score of 14.2% on AA-Omniscience Hallucination Rate.
How many models are evaluated on AA-Omniscience Hallucination Rate?
189 AI models have been evaluated on AA-Omniscience Hallucination Rate on BenchLM.
Compare top models on AA-Omniscience Hallucination Rate
Know when it’s worth switching models
The model to choose, the cheaper alternative, and the release we would wait on.
Read a sample issueJoin 2,000+ readers.
One email each week. Unsubscribe anytime.