# Grok 4.6 vs MiniMax M3

> Side-by-side benchmark comparison for Grok 4.6 and MiniMax M3 across agentic, coding, multimodal, knowledge, reasoning, multilingual, and math tasks.

- Canonical page: https://benchlm.ai/compare/grok-4-6-vs-minimax-m3
- Last updated: September 25, 2026

- Shared sourced benchmarks: 5
- HTML indexing: indexable
- Ranking lane: BenchAlign v5.7

## Quick Verdict

Pick Grok 4.6 if you want the stronger benchmark profile. MiniMax M3 only makes more sense when its price, context window, or workload-specific category wins matter more than the overall score.

## Summary

- Grok 4.6 leads overall 69.19 to 54.86.
- The clearest category separation is in agentic, where the averages are 67.9 for Grok 4.6 and 39.9 for MiniMax M3.
- The biggest single benchmark swing is Terminal-Bench 2.1 (Vals) in Agentic, with scores of 78.3% and 53.6%.
- MiniMax M3 is the cheaper option on output tokens, which matters if you expect large responses or heavy interactive use.
- MiniMax M3 also has the larger context window at 1M.

## Model Snapshot

| Property | Grok 4.6 | MiniMax M3 |
|----------|----------|----------|
| Creator | xAI | MiniMax |
| Type | Proprietary | Open Weight |
| Reasoning | Reasoning | Non-Reasoning |
| Context | 500K | 1M |
| Overall Score | 69.19 | 54.86 |
| Benchmarks Covered | 17 | 27 |
| Pricing (input/output) | $2.00 / $6.00 | $0.30 / $1.20 |

## Category Breakdown

### Agentic

- Winner: Grok 4.6
- Grok 4.6 public-lane score: 67.9 (Supported · #8/105)
- MiniMax M3 public-lane score: 39.9 (Supported · #45/105)

| Benchmark | Grok 4.6 | MiniMax M3 | Winner |
|-----------|-----------|-----------|--------|
| Terminal-Bench 3.0 | 26.5% | Coming soon | Coming soon |
| APEX-Agents | 57.5% | Coming soon | Coming soon |
| Terminal-Bench 2.1 (Vals) | 78.3% | 53.6% | Grok 4.6 |
| ApprenticeBench | 13% | Coming soon | Coming soon |
| Terminal-Bench 2.1 | Coming soon | 66.0% | Coming soon |
| BrowseComp | Coming soon | 83.5% | Coming soon |
| OSWorld-Verified | Coming soon | 70.1% | Coming soon |
| MCP Atlas | Coming soon | 74.2% | Coming soon |
| Claw-Eval | Coming soon | 74.5% | Coming soon |
| BankerToolBench | Coming soon | 76.1% | Coming soon |
| ResearchClawBench | Coming soon | 19.8% | Coming soon |
| OSWorld 2.0 | Coming soon | 4.6% | Coming soon |

### Coding

- Winner: Grok 4.6
- Grok 4.6 public-lane score: 62 (Supported · #14/135)
- MiniMax M3 public-lane score: 39.2 (Supported · #59/135)

| Benchmark | Grok 4.6 | MiniMax M3 | Winner |
|-----------|-----------|-----------|--------|
| Bug Hunt Bench | 27 fixes | Coming soon | Coming soon |
| DeepSWE | 65.9% | Coming soon | Coming soon |
| cursorBench32 | 70.8% | Coming soon | Coming soon |
| FrontierCode 1.1 Extended | 61.3% | Coming soon | Coming soon |
| VulcanBench v3 | 87.0% | Coming soon | Coming soon |
| FrontierSWE v2 | 25.3% | Coming soon | Coming soon |
| LiveCodeBench (Vals) | 88.2% | 82.2% | Grok 4.6 |
| SWE-bench (Vals) | 95.6% | 75.0% | Grok 4.6 |
| SWE-bench Verified | Coming soon | 80.5% | Coming soon |
| SWE-bench Pro | Coming soon | 59% | Coming soon |
| Terminal-Bench 2.1 | Coming soon | 66.0% | Coming soon |
| NL2Repo | Coming soon | 42.1% | Coming soon |
| VIBE V2 | Coming soon | 50.1% | Coming soon |
| SVG-Bench | Coming soon | 63.7% | Coming soon |
| KernelBench Hard | Coming soon | 28.8% | Coming soon |
| OpenHarmony Bench | Coming soon | 48.4% | Coming soon |

### Reasoning

- Winner: Not comparable
- Grok 4.6 public-lane score: 57.6 (#16/19)
- MiniMax M3 public-lane score: 79.1 (Unranked · 2 rankable rows)

| Benchmark | Grok 4.6 | MiniMax M3 | Winner |
|-----------|-----------|-----------|--------|
| ARC-AGI-1 | 87.00% | Coming soon | Coming soon |
| ARC-AGI-2 | 67.1% | Coming soon | Coming soon |
| ARC-AGI-3 | 2.1% | Coming soon | Coming soon |

### Multimodal & Grounded

- Winner: Not comparable
- Grok 4.6 public-lane score: Coming soon
- MiniMax M3 public-lane score: 52 (#36/50)

| Benchmark | Grok 4.6 | MiniMax M3 | Winner |
|-----------|-----------|-----------|--------|
| OfficeQA Pro | Coming soon | 45.1% | Coming soon |
| OmniDocBench 1.5 | Coming soon | 91.6% | Coming soon |
| MMMU-Pro | Coming soon | 78.1% | Coming soon |
| VideoMMMU | Coming soon | 84.6% | Coming soon |
| Video-MME (with subtitle) | Coming soon | 85.4% | Coming soon |

### Knowledge

- Winner: Grok 4.6
- Grok 4.6 public-lane score: 69 (Supported · #13/158)
- MiniMax M3 public-lane score: 49.3 (Supported · #59/158)

| Benchmark | Grok 4.6 | MiniMax M3 | Winner |
|-----------|-----------|-----------|--------|
| GPQA Diamond (Vals) | 94.7% | 92.7% | Grok 4.6 |
| MMLU-Pro (Vals) | 89.4% | 84.2% | Grok 4.6 |

### Instruction Following

- Winner: Not comparable
- Grok 4.6 public-lane score: Coming soon
- MiniMax M3 public-lane score: 92.4 (#5/124)

### Mathematics

- Winner: Not comparable
- Grok 4.6 public-lane score: Coming soon
- MiniMax M3 public-lane score: Coming soon

| Benchmark | Grok 4.6 | MiniMax M3 | Winner |
|-----------|-----------|-----------|--------|
| USAMO 2026 | Coming soon | 85.7% | Coming soon |

## FAQ

### Which is better overall, Grok 4.6 or MiniMax M3?

Grok 4.6 is ahead overall on BenchLM right now.

### Where is the biggest gap between Grok 4.6 and MiniMax M3?

The widest category gap is in agentic, where the averages are 67.9 for Grok 4.6 and 39.9 for MiniMax M3.

### How many benchmarks does Grok 4.6 cover on BenchLM?

Grok 4.6 currently has 17 sourced benchmark scores on BenchLM.

### How many benchmarks does MiniMax M3 cover on BenchLM?

MiniMax M3 currently has 27 sourced benchmark scores on BenchLM.

## Related Comparisons

- [Grok 4.6 vs GPT-6 Astra](/compare/gpt-6-astra-vs-grok-4-6)
- [MiniMax M3 vs GPT-6 Astra](/compare/gpt-6-astra-vs-minimax-m3)
- [Grok 4.6 vs Claude Opus 5.5](/compare/claude-opus-5-5-vs-grok-4-6)
- [MiniMax M3 vs Claude Opus 5.5](/compare/claude-opus-5-5-vs-minimax-m3)

## Explore More

- [Grok 4.6 profile](/models/grok-4-6)
- [MiniMax M3 profile](/models/minimax-m3)
- [Compare Pricing](/llm-pricing)
- [Alternative Finder](/tools/alternative-finder)
- [LLM Selector](/tools/llm-selector)
- [Overall Rankings](/best/overall)
