OmniDocBench 1.5
We show this table for reference; we do not rank on it.
A document understanding benchmark used in frontier-model comparison tables to measure extraction and grounded reasoning quality on complex documents.
Benchmark score on OmniDocBench 1.5 — September 27, 2026
We compile the OmniDocBench 1.5 rows from provider self-reports. Qwen3.8 Max leads the table at 92.1%, followed by MiniMax M3 (91.6%) and Qwen3.7 Plus (91.4%). We do not use these results to rank models overall.
Qwen3.8 Max
Alibaba
MiniMax M3
MiniMax
Qwen3.7 Plus
Alibaba
6 modelsMultimodal & GroundedCurrentDisplay onlyUpdated September 27, 2026
Benchmark score table (6 models)
ScoreAmong the reported OmniDocBench 1.5 rows, Qwen3.8 Max is first at 92.1%. The third row is 0.7 points behind. The broader top-10 range is 16.3 points, so the table still separates the published systems.
6 models have been evaluated on OmniDocBench 1.5. The benchmark falls in the Multimodal & Grounded category. OmniDocBench 1.5 is currently displayed for reference but excluded from the scoring formula, so it does not directly affect overall rankings.
About OmniDocBench 1.5
Year
2026
Tasks
Document understanding tasks
Format
Document understanding benchmark
Difficulty
Grounded document reasoning
BenchLM stores OmniDocBench 1.5 as the higher-is-better score format used in current first-party comparison tables. Earlier low-is-better error-style rows are intentionally not mixed into this benchmark key.
Freshness and provenance
Version
OmniDocBench 1.5 2026
Refresh cadence
Quarterly
Staleness state
Current
Question availability
Public benchmark set
BenchLM uses freshness metadata to decide whether a benchmark should still be treated as a strong differentiator, a benchmark to watch, or a display-only reference. For the full scoring policy, see the BenchLM methodology page.
Questions
What does OmniDocBench 1.5 measure?
A document understanding benchmark used in frontier-model comparison tables to measure extraction and grounded reasoning quality on complex documents.
Which model scores highest on OmniDocBench 1.5?
Qwen3.8 Max by Alibaba currently leads with a score of 92.1% on OmniDocBench 1.5.
How many models are evaluated on OmniDocBench 1.5?
6 AI models have been evaluated on OmniDocBench 1.5 on BenchLM.
Compare top models on OmniDocBench 1.5
Know when it’s worth switching models
The model to choose, the cheaper alternative, and the release we would wait on.
Read a sample issueJoin 2,000+ readers.
One email each week. Unsubscribe anytime.