# Claude Fable 5.1 vs Grok 4.7

> Side-by-side benchmark comparison for Claude Fable 5.1 and Grok 4.7 across agentic, coding, multimodal, knowledge, reasoning, multilingual, and math tasks.

- Canonical page: https://benchlm.ai/compare/claude-fable-5-1-vs-grok-4-7
- Last updated: September 21, 2026

- Shared sourced benchmarks: 4
- HTML indexing: indexable
- Ranking lane: BenchAlign v5

## Quick Verdict

Pick Claude Fable 5.1 if you want the stronger benchmark profile. Grok 4.7 only makes more sense when its price, context window, or workload-specific category wins matter more than the overall score.

## Summary

- Claude Fable 5.1 leads overall 84.63 to 68.
- The clearest category separation is in reasoning, where the averages are 79.4 for Claude Fable 5.1 and 73.7 for Grok 4.7.
- The biggest single benchmark swing is Terminal-Bench 4.0 in Agentic, with scores of 55.80% and 38.00%.
- Grok 4.7 is the cheaper option on output tokens, which matters if you expect large responses or heavy interactive use.
- Claude Fable 5.1 also has the larger context window at 1M.

## Model Snapshot

| Property | Claude Fable 5.1 | Grok 4.7 |
|----------|----------|----------|
| Creator | Anthropic | xAI |
| Type | Proprietary | Proprietary |
| Reasoning | Reasoning | Reasoning |
| Context | 1M | 500K |
| Overall Score | 84.63 | null |
| Benchmarks Covered | 26 | 6 |
| Pricing (input/output) | $10.00 / $50.00 | $2.00 / $6.00 |

## Category Breakdown

### Agentic

- Winner: Coming soon
- Claude Fable 5.1 public-lane score: 80.3 (Supported · #1/154)
- Grok 4.7 public-lane score: Coming soon

| Benchmark | Claude Fable 5.1 | Grok 4.7 | Winner |
|-----------|-----------|-----------|--------|
| Terminal-Bench 4.0 | 55.80% | 38.00% | Claude Fable 5.1 |
| Terminal-Bench-Science 0.1 | 52.6% | Coming soon | Coming soon |
| OSWorld 2.0 | 41.7% | Coming soon | Coming soon |
| AutomationBench | 31.4% | Coming soon | Coming soon |
| Toolathlon-Verified | 77.8% | Coming soon | Coming soon |
| Toolathlon Verified Pass@3 | 81.5% | Coming soon | Coming soon |
| Toolathlon Verified Pass³ | 73.1% | Coming soon | Coming soon |
| Toolathlon Verified avg. turns | 23.7 turns | Coming soon | Coming soon |
| Terminal-Bench 2.1 (Vals) | 85.0% | 76.0% | Claude Fable 5.1 |
| ApprenticeBench | 72% | Coming soon | Coming soon |

### Coding

- Winner: Coming soon
- Claude Fable 5.1 public-lane score: 84.5 (Supported · #1/156)
- Grok 4.7 public-lane score: Coming soon

| Benchmark | Claude Fable 5.1 | Grok 4.7 | Winner |
|-----------|-----------|-----------|--------|
| Bug Hunt Bench | 43 fixes | Coming soon | Coming soon |
| SWE-bench Pro | 81.2% | Coming soon | Coming soon |
| SWE Multilingual | 89.1% | Coming soon | Coming soon |
| SWE Multimodal | 54.7% | Coming soon | Coming soon |
| DeepSWE | 67.4% | 71.0% | Grok 4.7 |
| FrontierSWE v2 | 56.3% | Coming soon | Coming soon |
| ProgramBench | 87.6% | Coming soon | Coming soon |
| cursorBench32 | 73.4% | Coming soon | Coming soon |
| LiveCodeBench (Vals) | 90.5% | Coming soon | Coming soon |
| cursorBench40 | 51.8% | 46.3% | Claude Fable 5.1 |
| EEBench | Coming soon | 64.0% | Coming soon |

### Reasoning

- Winner: Coming soon
- Claude Fable 5.1 public-lane score: 79.4 (#2/19)
- Grok 4.7 public-lane score: 73.7 (Unranked · 2 rankable rows)

| Benchmark | Claude Fable 5.1 | Grok 4.7 | Winner |
|-----------|-----------|-----------|--------|
| ARC-AGI-1 | 97.50% | Coming soon | Coming soon |
| ARC-AGI-2 | 90% | Coming soon | Coming soon |

### Knowledge

- Winner: Coming soon
- Claude Fable 5.1 public-lane score: 87 (Supported · #1/186)
- Grok 4.7 public-lane score: Coming soon

| Benchmark | Claude Fable 5.1 | Grok 4.7 | Winner |
|-----------|-----------|-----------|--------|
| HLE | 65% | Coming soon | Coming soon |
| HLE w/o tools | 60.9% | Coming soon | Coming soon |
| GPQA Diamond (Vals) | 93.4% | Coming soon | Coming soon |
| MMLU-Pro (Vals) | 92.4% | Coming soon | Coming soon |
| HealthBench Professional | Coming soon | 56.7% | Coming soon |

## FAQ

### Which is better overall, Claude Fable 5.1 or Grok 4.7?

Claude Fable 5.1 is ahead overall on BenchLM right now.

### Where is the biggest gap between Claude Fable 5.1 and Grok 4.7?

The widest category gap is in reasoning, where the averages are 79.4 for Claude Fable 5.1 and 73.7 for Grok 4.7.

### How many benchmarks does Claude Fable 5.1 cover on BenchLM?

Claude Fable 5.1 currently has 26 sourced benchmark scores on BenchLM.

### How many benchmarks does Grok 4.7 cover on BenchLM?

Grok 4.7 currently has 6 sourced benchmark scores on BenchLM.

## Related Comparisons

- [Claude Fable 5.1 vs Claude Fable 5](/compare/claude-fable-vs-claude-fable-5-1)
- [Grok 4.7 vs Claude Fable 5](/compare/claude-fable-vs-grok-4-7)
- [Claude Fable 5.1 vs GPT-6 Astra](/compare/claude-fable-5-1-vs-gpt-6-astra)
- [Grok 4.7 vs GPT-6 Astra](/compare/gpt-6-astra-vs-grok-4-7)

## Explore More

- [Claude Fable 5.1 profile](/models/claude-fable-5-1)
- [Grok 4.7 profile](/models/grok-4-7)
- [Compare Pricing](/llm-pricing)
- [Alternative Finder](/tools/alternative-finder)
- [LLM Selector](/tools/llm-selector)
- [Overall Rankings](/best/overall)
