# Claude Opus 4.5 Thinking Benchmark Scores & Performance

> Claude Opus 4.5 Thinking by Anthropic scores 56.09/100 overall, ranking #96 out of 411 AI models.

Canonical page: https://benchlm.ai/models/claude-opus-4-5-thinking

Last updated: September 4, 2026

## Model Details

| Property | Value |
|----------|-------|
| Creator | Anthropic |
| Source Type | Proprietary |
| Reasoning Type | Reasoning |
| Context Window | 200K |
| Overall Score | 56.09/100 |
| Overall Rank | #96 of 411 |

## Family & Coverage

- Family: Claude Opus 4.5
- Variant: reasoning (thinking)
- Benchmarks covered: 15 of 422
- Sibling models: [Claude Opus 4.5](/models/claude-opus-4-5)
- Coverage note: BenchLM currently has partial benchmark coverage for this model, so the overall score is conservative.

## Agentic Benchmarks

| Benchmark | Score |
|-----------|-------|
| [τ²-bench results](/benchmarks/tau2-bench) | 89.5% |

## Coding Benchmarks

| Benchmark | Score |
|-----------|-------|
| [Vibe Code Bench](/benchmarks/vibecodebench) | 20.63% |
| [AA-SciCode](/benchmarks/aascicode) | 49.5% |

## Multimodal & Grounded Benchmarks

| Benchmark | Score |
|-----------|-------|
| [AA-MMMU-Pro](/benchmarks/aammmupro) | 74.0% |
| [Design Arena Website](/benchmarks/designarenawebsite) | 1262 |

## Reasoning Benchmarks

| Benchmark | Score |
|-----------|-------|
| [AA-LCR](/benchmarks/lcr) | 76.0% |
| [CritPt](/benchmarks/critpt) | 4.6% |

## Knowledge Benchmarks

| Benchmark | Score |
|-----------|-------|
| [Artificial Analysis Intelligence Index](/benchmarks/artificialanalysis) | 41.9% |
| [AA-GPQA Diamond](/benchmarks/aagpqadiamond) | 86.6% |
| [AA-HLE](/benchmarks/aahle) | 30.1% |
| [AA-Omniscience Index](/benchmarks/aaomniscienceindex) | 14.0% |
| [AA-Omniscience Accuracy](/benchmarks/omniscienceaccuracy) | 46.6% |
| [AA-Omniscience Hallucination Rate](/benchmarks/omnisciencehallucinationrate) | 61.0% |
| [AA MMLU-Pro](/benchmarks/aammlupro) | 89.5% |

## Instruction Following Benchmarks

| Benchmark | Score |
|-----------|-------|
| [AA-IFBench](/benchmarks/aaifbench) | 58.0% |

## Other Anthropic Models

- [Claude Fable 5.1](/models/claude-fable-5-1) - Score: 82.95
- [Claude Fable 5](/models/claude-fable) - Score: 80.9
- [Claude Opus 5](/models/claude-opus-5) - Score: 80.66
- [Claude Opus 4.8](/models/claude-opus-4-8) - Score: 72.3
- [Claude Sonnet 5](/models/claude-sonnet-5) - Score: 70.76
- [Claude Opus 4.7](/models/claude-opus-4-7) - Score: 70.22
- [Claude Opus 4.7 (Adaptive)](/models/claude-opus-4-7-adaptive) - Score: 69.94
- [Claude Opus 4.6](/models/claude-opus-4-6) - Score: 69.84
- [Claude Sonnet 4.6](/models/claude-sonnet-4-6) - Score: 64.21
- [Claude Opus 4.6 (Adaptive)](/models/claude-opus-4-6-thinking) - Score: 63.58
- [Claude Opus 4.5](/models/claude-opus-4-5) - Score: 60.24
- [Claude Sonnet 4.5](/models/claude-sonnet-4-5) - Score: 53.74
- [Claude Haiku 4.5](/models/claude-haiku-4-5) - Score: 52.73
- [Claude 4.1 Opus](/models/claude-4-1-opus) - Score: 44.39
- [Claude 4 Sonnet](/models/claude-4-sonnet) - Score: 41.76
- [Claude 3 Opus](/models/claude-3-opus) - Score: 36.84
- [Claude 4.1 Opus Thinking](/models/claude-4-1-opus-thinking) - Score: 34.62
- [Claude 3.5 Sonnet](/models/claude-3-5-sonnet) - Score: 31.93
- [Claude 3 Haiku](/models/claude-3-haiku) - Score: 17.34
- [Claude Mythos 5](/models/claude-mythos-5) - Score: not computed
- [Claude Mythos 5.1](/models/claude-mythos-5-1) - Score: not computed
- [Claude Mythos Preview](/models/claude-mythos-preview) - Score: not computed
- [Claude Haiku 4.5 Thinking](/models/claude-haiku-4-5-thinking) - Score: not computed
- [Claude Sonnet 4.5 Thinking](/models/claude-sonnet-4-5-thinking) - Score: not computed
