OmniDocBench v1.6 (OmniDocBench 1.6)
We show this table for reference; we do not rank on it.
A document parsing benchmark whose overall score composites edit distance, CDM and TEDS sub-scores across diverse PDF layouts.
Benchmark score on OmniDocBench 1.6 — September 27, 2026
We compile the OmniDocBench 1.6 rows from provider self-reports. Ternary Bonsai 2 27B leads the table at 89.1%. We do not use these results to rank models overall.
1 modelMultimodal & GroundedRefreshingDisplay onlyUpdated September 27, 2026
Benchmark score table (1 model)
ScoreAbout OmniDocBench 1.6
Year
2024
Tasks
Document parsing and extraction
Format
Composite of edit distance, CDM and TEDS
Difficulty
Grounded document reasoning
BenchLM stores OmniDocBench v1.6 on its own higher-is-better key. The v1.5 lane is kept separate because the releases differ in both annotation set and composite definition, and earlier low-is-better error-style rows are not mixed into either key.
Freshness and provenance
Version
OmniDocBench 1.6 2024
Refresh cadence
Annual
Staleness state
Refreshing
Question availability
Public benchmark set
BenchLM uses freshness metadata to decide whether a benchmark should still be treated as a strong differentiator, a benchmark to watch, or a display-only reference. For the full scoring policy, see the BenchLM methodology page.
Questions
What does OmniDocBench 1.6 measure?
A document parsing benchmark whose overall score composites edit distance, CDM and TEDS sub-scores across diverse PDF layouts.
Which model scores highest on OmniDocBench 1.6?
Ternary Bonsai 2 27B by Prism ML currently leads with a score of 89.1% on OmniDocBench 1.6.
How many models are evaluated on OmniDocBench 1.6?
1 AI models have been evaluated on OmniDocBench 1.6 on BenchLM.
Know when it’s worth switching models
The model to choose, the cheaper alternative, and the release we would wait on.
Read a sample issueJoin 2,000+ readers.
One email each week. Unsubscribe anytime.