# Claude Haiku 4.5 vs Grok 4.6

> Side-by-side benchmark comparison for Claude Haiku 4.5 and Grok 4.6 across agentic, coding, multimodal, knowledge, reasoning, multilingual, and math tasks.

- Canonical page: https://benchlm.ai/compare/claude-haiku-4-5-vs-grok-4-6
- Last updated: September 25, 2026

- Shared sourced benchmarks: 6
- HTML indexing: indexable
- Ranking lane: BenchAlign v5.7

## Quick Verdict

Pick Grok 4.6 if you want the stronger benchmark profile. Claude Haiku 4.5 only makes more sense when its price, context window, or workload-specific category wins matter more than the overall score.

## Summary

- Grok 4.6 leads overall 69.19 to 42.48.
- The clearest category separation is in agentic, where the averages are 21.7 for Claude Haiku 4.5 and 67.9 for Grok 4.6.
- The biggest single benchmark swing is LiveCodeBench (Vals) in Coding, with scores of 41.2% and 88.2%.
- Claude Haiku 4.5 is the cheaper option on output tokens, which matters if you expect large responses or heavy interactive use.
- Grok 4.6 also has the larger context window at 500K.

## Model Snapshot

| Property | Claude Haiku 4.5 | Grok 4.6 |
|----------|----------|----------|
| Creator | Anthropic | xAI |
| Type | Proprietary | Proprietary |
| Reasoning | Non-Reasoning | Reasoning |
| Context | 200K | 500K |
| Overall Score | 42.48 | 69.19 |
| Benchmarks Covered | 10 | 17 |
| Pricing (input/output) | $1.00 / $5.00 | $2.00 / $6.00 |

## Category Breakdown

### Agentic

- Winner: Grok 4.6
- Claude Haiku 4.5 public-lane score: 21.7 (Supported · #85/105)
- Grok 4.6 public-lane score: 67.9 (Supported · #8/105)

| Benchmark | Claude Haiku 4.5 | Grok 4.6 | Winner |
|-----------|-----------|-----------|--------|
| JobBench | 16.0% | Coming soon | Coming soon |
| Terminal-Bench 2.1 (Vals) | 43.8% | 78.3% | Grok 4.6 |
| Terminal-Bench 3.0 | Coming soon | 26.5% | Coming soon |
| APEX-Agents | Coming soon | 57.5% | Coming soon |
| ApprenticeBench | Coming soon | 13% | Coming soon |

### Coding

- Winner: Grok 4.6
- Claude Haiku 4.5 public-lane score: 19.8 (Supported · #117/135)
- Grok 4.6 public-lane score: 62 (Supported · #14/135)

| Benchmark | Claude Haiku 4.5 | Grok 4.6 | Winner |
|-----------|-----------|-----------|--------|
| SWE-bench Verified | 73.3% | Coming soon | Coming soon |
| VulcanBench v3 | 76.2% | 87.0% | Grok 4.6 |
| LiveCodeBench (Vals) | 41.2% | 88.2% | Grok 4.6 |
| SWE-bench (Vals) | 66.6% | 95.6% | Grok 4.6 |
| Bug Hunt Bench | Coming soon | 27 fixes | Coming soon |
| DeepSWE | Coming soon | 65.9% | Coming soon |
| cursorBench32 | Coming soon | 70.8% | Coming soon |
| FrontierCode 1.1 Extended | Coming soon | 61.3% | Coming soon |
| FrontierSWE v2 | Coming soon | 25.3% | Coming soon |

### Reasoning

- Winner: Not comparable
- Claude Haiku 4.5 public-lane score: Coming soon
- Grok 4.6 public-lane score: 57.6 (#16/19)

| Benchmark | Claude Haiku 4.5 | Grok 4.6 | Winner |
|-----------|-----------|-----------|--------|
| ARC-AGI-1 | Coming soon | 87.00% | Coming soon |
| ARC-AGI-2 | Coming soon | 67.1% | Coming soon |
| ARC-AGI-3 | Coming soon | 2.1% | Coming soon |

### Knowledge

- Winner: Directional only
- Claude Haiku 4.5 public-lane score: 34.5 (Estimated · #111/158)
- Grok 4.6 public-lane score: 69 (Supported · #13/158)

| Benchmark | Claude Haiku 4.5 | Grok 4.6 | Winner |
|-----------|-----------|-----------|--------|
| GPQA Diamond (Vals) | 72.2% | 94.7% | Grok 4.6 |
| MMLU-Pro (Vals) | 78.7% | 89.4% | Grok 4.6 |

### Mathematics

- Winner: Not comparable
- Claude Haiku 4.5 public-lane score: 28.8 (Unranked · 2 rankable rows)
- Grok 4.6 public-lane score: Coming soon

| Benchmark | Claude Haiku 4.5 | Grok 4.6 | Winner |
|-----------|-----------|-----------|--------|
| FrontierMath v2 (Tiers 1-3) | 5.903% | Coming soon | Coming soon |
| FrontierMath v2 (Tier 4) | 2.083% | Coming soon | Coming soon |

## FAQ

### Which is better overall, Claude Haiku 4.5 or Grok 4.6?

Grok 4.6 is ahead overall on BenchLM right now.

### Where is the biggest gap between Claude Haiku 4.5 and Grok 4.6?

The widest category gap is in agentic, where the averages are 21.7 for Claude Haiku 4.5 and 67.9 for Grok 4.6.

### How many benchmarks does Claude Haiku 4.5 cover on BenchLM?

Claude Haiku 4.5 currently has 10 sourced benchmark scores on BenchLM.

### How many benchmarks does Grok 4.6 cover on BenchLM?

Grok 4.6 currently has 17 sourced benchmark scores on BenchLM.

## Related Comparisons

- [Claude Haiku 4.5 vs Claude Haiku 4.5 Thinking](/compare/claude-haiku-4-5-vs-claude-haiku-4-5-thinking)
- [Grok 4.6 vs Claude Haiku 4.5 Thinking](/compare/claude-haiku-4-5-thinking-vs-grok-4-6)
- [Claude Haiku 4.5 vs GPT-6 Astra](/compare/claude-haiku-4-5-vs-gpt-6-astra)
- [Grok 4.6 vs GPT-6 Astra](/compare/gpt-6-astra-vs-grok-4-6)

## Explore More

- [Claude Haiku 4.5 profile](/models/claude-haiku-4-5)
- [Grok 4.6 profile](/models/grok-4-6)
- [Compare Pricing](/llm-pricing)
- [Alternative Finder](/tools/alternative-finder)
- [LLM Selector](/tools/llm-selector)
- [Overall Rankings](/best/overall)
