# Claude Fable 5.1 Benchmark Scores & Performance

> Claude Fable 5.1 by Anthropic scores 83.34/100 overall, ranking #3 out of 505 AI models.

Canonical page: https://benchlm.ai/models/claude-fable-5-1

Last updated: September 22, 2026

## Model Details

| Property | Value |
|----------|-------|
| Creator | Anthropic |
| Source Type | Proprietary |
| Reasoning Type | Reasoning |
| Context Window | 1M |
| Official model card | [Claude Fable 5.1 and Claude Mythos 5.1 system card](https://www-cdn.anthropic.com/0339e6a7c5c7b87f5c07798616dc32c215d14235/Claude%20Fable%205.1%20&%20Claude%20Mythos%205.1%20System%20Card.pdf) |
| Overall Score | 83.34/100 |
| Overall Rank | #3 of 505 |

## Family & Coverage

- Family: Claude Fable
- Variant: base
- Benchmarks covered: 47 of 481
- Sibling models: [Claude Fable 5](/models/claude-fable)
- Related earlier model: [Claude Fable 5](/models/claude-fable)
- Coverage note: BenchLM currently has partial benchmark coverage for this model, so the overall score is conservative.

## Agentic Benchmarks

| Benchmark | Score |
|-----------|-------|
| [Terminal-Bench 4.0](/benchmarks/terminal-bench-4) | 55.80% |
| [Terminal-Bench-Science 0.1](/benchmarks/terminal-bench-science) | 52.6% |
| [OSWorld 2.0](/benchmarks/osworld2) | 41.7% |
| [AutomationBench](/benchmarks/automationbench) | 31.4% |
| [Toolathlon-Verified](/benchmarks/toolathlonverified) | 77.8% |
| [Toolathlon Verified Pass@3](/benchmarks/toolathlonverifiedpass3) | 81.5% |
| [Toolathlon Verified Pass³](/benchmarks/toolathlonverifiedpass3all) | 73.1% |
| [Toolathlon Verified avg. turns](/benchmarks/toolathlonverifiedavgturns) | 23.7 turns |
| [GDPval-AA](/benchmarks/gdpvalaanormalized) | 61.7% |
| [AA Briefcase](/benchmarks/aabriefcaseelo) | 1678 |
| [AA Harvey LAB](/benchmarks/aaharveylab) | 93.0% |
| [AA Tau3 Banking](/benchmarks/aatau3banking) | 47.2% |
| [Terminal-Bench 2.1 (Vals)](/benchmarks/valsterminalbench21) | 85.0% |
| [AA AutomationBench](/benchmarks/aaautomationbench) | 59.4% |
| [GDPval-AA](/benchmarks/gdpvalaa) | 1735 |
| [AA Agentic Index](/benchmarks/aaagenticindex) | 58.0% |
| [GDP.pdf](/benchmarks/aagdppdf) | 26.2% |
| [AA-AnalystAgent](/benchmarks/aaanalystagent) | 57.5% |
| [ApprenticeBench](/benchmarks/apprenticebench) | 72% |

## Coding Benchmarks

| Benchmark | Score |
|-----------|-------|
| [Bug Hunt Bench](/benchmarks/bug-hunt-bench) | 43 fixes |
| [SWE-bench Pro](/benchmarks/swe-bench-pro) | 81.2% |
| [SWE Multilingual](/benchmarks/swe-bench-multilingual) | 89.1% |
| [SWE Multimodal](/benchmarks/swe-bench-multimodal) | 54.7% |
| [DeepSWE](/benchmarks/deepswe) | 67.4% |
| [FrontierSWE v2](/benchmarks/frontierswev2) | 56.3% |
| [ProgramBench](/benchmarks/programbench) | 87.6% |
| [CursorBench 3.2](/benchmarks/cursorbench32) | 73.4% |
| [AA-SciCode](/benchmarks/aascicode) | 63.1% |
| [LiveCodeBench (Vals)](/benchmarks/valslivecodebench) | 90.5% |
| [AA Coding Index](/benchmarks/aacodingindex) | 81.6% |
| [CursorBench 4.0](/benchmarks/cursorbench40) | 51.8% |

## Multimodal & Grounded Benchmarks

| Benchmark | Score |
|-----------|-------|
| [Design Arena Website](/benchmarks/designarenawebsite) | 1320 |

## Reasoning Benchmarks

| Benchmark | Score |
|-----------|-------|
| [ARC-AGI-1](/benchmarks/arcagi1) | 97.50% |
| [ARC-AGI-2](/benchmarks/arc-agi-2) | 90% |
| [AA-LCR](/benchmarks/lcr) | 85.3% |
| [CritPt](/benchmarks/critpt) | 29.7% |
| [MLCR-AA](/benchmarks/aamlcr) | 71.1% |

## Knowledge Benchmarks

| Benchmark | Score |
|-----------|-------|
| [HLE](/benchmarks/hle) | 65% |
| [HLE w/o tools](/benchmarks/hlenotools) | 60.9% |
| [Artificial Analysis Intelligence Index](/benchmarks/artificialanalysis) | 53.4% |
| [AA-GPQA Diamond](/benchmarks/aagpqadiamond) | 93.7% |
| [AA-HLE](/benchmarks/aahle) | 59.1% |
| [AA-Omniscience Index](/benchmarks/aaomniscienceindex) | 43.5% |
| [AA-Omniscience Accuracy](/benchmarks/omniscienceaccuracy) | 67.2% |
| [AA-Omniscience Hallucination Rate](/benchmarks/omnisciencehallucinationrate) | 72.6% |
| [GPQA Diamond (Vals)](/benchmarks/valsgpqadiamond) | 93.4% |
| [MMLU-Pro (Vals)](/benchmarks/valsmmlupro) | 92.4% |

## Other Anthropic Models

- [Claude Opus 5.5](/models/claude-opus-5-5) - Score: 88.45
- [Claude Opus 5](/models/claude-opus-5) - Score: 80.35
- [Claude Fable 5](/models/claude-fable) - Score: 79.53
- [Claude Opus 4.8](/models/claude-opus-4-8) - Score: 70.86
- [Claude Opus 4.7 (Adaptive)](/models/claude-opus-4-7-adaptive) - Score: 68.83
- [Claude Sonnet 5](/models/claude-sonnet-5) - Score: 67.25
- [Claude Opus 4.7](/models/claude-opus-4-7) - Score: 66.35
- [Claude Opus 4.6](/models/claude-opus-4-6) - Score: 64.33
- [Claude Sonnet 4.6](/models/claude-sonnet-4-6) - Score: 56.73
- [Claude Opus 4.5](/models/claude-opus-4-5) - Score: 55.9
- [Claude Opus 4.5 Thinking](/models/claude-opus-4-5-thinking) - Score: 55.47
- [Claude Sonnet 4.5](/models/claude-sonnet-4-5) - Score: 47.86
- [Claude Haiku 4.5](/models/claude-haiku-4-5) - Score: 41.55
- [Claude 4.1 Opus](/models/claude-4-1-opus) - Score: 38.45
- [Claude 4.1 Opus Thinking](/models/claude-4-1-opus-thinking) - Score: 36.6
- [Claude 4 Sonnet](/models/claude-4-sonnet) - Score: 36.06
- [Claude 3.5 Sonnet](/models/claude-3-5-sonnet) - Score: 29.76
- [Claude 3 Opus](/models/claude-3-opus) - Score: 28.62
- [Claude 3 Haiku](/models/claude-3-haiku) - Score: 14.68
- [Claude Mythos 5](/models/claude-mythos-5) - Score: not computed
- [Claude Opus 4.6 (Adaptive)](/models/claude-opus-4-6-thinking) - Score: not computed
- [Claude Mythos 5.1](/models/claude-mythos-5-1) - Score: not computed
- [Claude Mythos Preview](/models/claude-mythos-preview) - Score: not computed
- [Claude Haiku 4.5 Thinking](/models/claude-haiku-4-5-thinking) - Score: not computed
- [Claude Sonnet 4.5 Thinking](/models/claude-sonnet-4-5-thinking) - Score: not computed
