# GDP.pdf mean criteria pass rate without tools (GDP.pdf (no tools))

> Professional document understanding over 100 real-world PDFs from ten domains.

Canonical page: https://benchlm.ai/benchmarks/gdppdf

- Category: [Multimodal & Grounded](/multimodal-grounded)
- Last updated: September 27, 2026

## About GDP.pdf (no tools)

- Year: 2026
- Tasks: 100 professional document prompts
- Format: Mean criteria pass rate
- Difficulty: Professional document reasoning
- Paper: [GDP.pdf](https://surgehq.ai/blog/gdp-pdf-can-100b-ai-models-master-the-documents-that-run-the-world)

Section 8.12.4 reports a five-run mean criteria pass rate from Anthropic's corrected internal harness without tools, using Opus 4.7 as judge.

GDP.pdf (no tools) is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Leaderboard (2 models)

| Rank | Model | Creator | Score |
|------|-------|---------|-------|
| 1 | [Claude Opus 5](/models/claude-opus-5) | Anthropic | 83.4% |
| 2 | [dots3-note Preview](/models/dots3-note-preview) | Dots Studio | 60.7% |

## FAQ

### What does GDP.pdf (no tools) measure?

Professional document understanding over 100 real-world PDFs from ten domains.

### Which model scores highest on GDP.pdf (no tools)?

Claude Opus 5 by Anthropic currently leads with a score of 83.4% on GDP.pdf (no tools).

### How many models are evaluated on GDP.pdf (no tools)?

2 AI models have been evaluated on GDP.pdf (no tools) on BenchLM.

### Does GDP.pdf (no tools) affect BenchLM's overall score?

Not directly. GDP.pdf (no tools) is still displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Compare Top Models on GDP.pdf (no tools)

- [Claude Opus 5 vs dots3-note Preview](/compare/claude-opus-5-vs-dots3-note-preview)
