# Qwen3.7 Max vs Qwen3.7 Plus

> Side-by-side benchmark comparison for Qwen3.7 Max and Qwen3.7 Plus across agentic, coding, multimodal, knowledge, reasoning, multilingual, and math tasks.

- Canonical page: https://benchlm.ai/compare/qwen3-7-max-vs-qwen3-7-plus
- Last updated: September 4, 2026

- Shared sourced benchmarks: 33
- HTML indexing: indexable
- Ranking lane: BenchAlign v5

## Quick Verdict

Pick Qwen3.7 Max if you want the stronger benchmark profile. Qwen3.7 Plus only makes more sense when its price, context window, or workload-specific category wins matter more than the overall score.

## Summary

- Qwen3.7 Max leads overall 68.56 to 62.29.
- The clearest category separation is in multilingual, where the averages are 100 for Qwen3.7 Max and 78.9 for Qwen3.7 Plus.
- The biggest single benchmark swing is Apex in Mathematics, with scores of 44.5% and 22.7%.

## Model Snapshot

| Property | Qwen3.7 Max | Qwen3.7 Plus |
|----------|----------|----------|
| Creator | Alibaba | Alibaba |
| Type | Proprietary | Proprietary |
| Reasoning | Reasoning | Reasoning |
| Context | 1M | 1M |
| Overall Score | 68.56 | 62.29 |
| Benchmarks Covered | 41 | 52 |
| Pricing (input/output) | Pricing unavailable / Pricing unavailable | N/A |

## Category Breakdown

### Agentic

- Winner: Qwen3.7 Max
- Qwen3.7 Max public-lane score: 42.8 (Supported · #110/151)
- Qwen3.7 Plus public-lane score: 37.7 (Supported · #126/151)

| Benchmark | Qwen3.7 Max | Qwen3.7 Plus | Winner |
|-----------|-----------|-----------|--------|
| Terminal-Bench 2.0 | 69.7% | 70.3% | Qwen3.7 Plus |
| QwenClawBench | 64.3% | 61.8% | Qwen3.7 Max |
| Claw-Eval | 65.2% | 62.7% | Qwen3.7 Max |
| BFCL v4 | 75.0% | 72.9% | Qwen3.7 Max |
| MCP Atlas | 76.4% | 73.2% | Qwen3.7 Max |
| VITA-Bench | 47.9% | 45.6% | Qwen3.7 Max |
| HLE w/ tools | 53.5% | Coming soon | Coming soon |
| Gert Labs | 64.27% | Coming soon | Coming soon |
| ResearchClawBench | 18.7% | Coming soon | Coming soon |
| Terminal-Bench 2.1 (Vals) | 61.0% | 52.8% | Qwen3.7 Max |
| DeepPlanning | Coming soon | 62.3% | Coming soon |
| OSWorld-Verified | Coming soon | 73.3% | Coming soon |
| AndroidWorld | Coming soon | 81.0% | Coming soon |
| OSWorld 2.0 | Coming soon | 2.8% | Coming soon |

### Coding

- Winner: Coming soon
- Qwen3.7 Max public-lane score: 49.5 (Supported · #77/183)
- Qwen3.7 Plus public-lane score: 53.1 (Estimated · #51/183)

| Benchmark | Qwen3.7 Max | Qwen3.7 Plus | Winner |
|-----------|-----------|-----------|--------|
| SWE-bench Verified | 80.4% | 77.7% | Qwen3.7 Max |
| SWE-bench Pro | 60.6% | 57.6% | Qwen3.7 Max |
| SWE Multilingual | 78.3% | 75.8% | Qwen3.7 Max |
| NL2Repo | 47.2% | 41.1% | Qwen3.7 Max |
| SciCode | 53.5% | 51.3% | Qwen3.7 Max |
| LiveCodeBench | 91.6% | 89.6% | Qwen3.7 Max |
| Terminal-Bench 2.0 | 69.7% | 70.3% | Qwen3.7 Plus |
| OpenHarmony Bench | 53.4% | Coming soon | Coming soon |
| LiveCodeBench (Vals) | 87.1% | Coming soon | Coming soon |
| SWE-bench (Vals) | 68.8% | Coming soon | Coming soon |

### Multimodal & Grounded

- Winner: Coming soon
- Qwen3.7 Max public-lane score: Coming soon
- Qwen3.7 Plus public-lane score: 72.5 (#18/48)

| Benchmark | Qwen3.7 Max | Qwen3.7 Plus | Winner |
|-----------|-----------|-----------|--------|
| MMMU-Pro | Coming soon | 79% | Coming soon |
| MathVision | Coming soon | 90.3% | Coming soon |
| CharXiv | Coming soon | 85.9% | Coming soon |
| ERQA | Coming soon | 69.8% | Coming soon |
| MedXpertQA (MM) | Coming soon | 71.0% | Coming soon |
| ScreenSpot Pro | Coming soon | 79.0% | Coming soon |
| SimpleVQA | Coming soon | 81.7% | Coming soon |
| MMSearch-Plus | Coming soon | 41.4% | Coming soon |
| RealWorldQA | Coming soon | 86.9% | Coming soon |
| OmniDocBench 1.5 | Coming soon | 91.4% | Coming soon |
| OCRBench V2 | Coming soon | 70.7% | Coming soon |
| ODINW13 | Coming soon | 51.1% | Coming soon |
| Video-MME (with subtitle) | Coming soon | 88.0% | Coming soon |
| VideoMMMU | Coming soon | 85.4% | Coming soon |
| MLVU (M-Avg) | Coming soon | 87.4% | Coming soon |

### Reasoning

- Winner: Coming soon
- Qwen3.7 Max public-lane score: 74.8 (Unranked · 3 rankable rows)
- Qwen3.7 Plus public-lane score: 73.7 (Unranked · 3 rankable rows)

| Benchmark | Qwen3.7 Max | Qwen3.7 Plus | Winner |
|-----------|-----------|-----------|--------|
| MRCRv2 | 90.4% | 91.7% | Qwen3.7 Plus |
| CritPt | 13.4% | 9.1% | Qwen3.7 Max |

### Knowledge

- Winner: Coming soon
- Qwen3.7 Max public-lane score: 62.4 (Supported · #28/181)
- Qwen3.7 Plus public-lane score: 56.8 (Estimated · #50/181)

| Benchmark | Qwen3.7 Max | Qwen3.7 Plus | Winner |
|-----------|-----------|-----------|--------|
| GPQA | 92.4% | 90.3% | Qwen3.7 Max |
| GPQA-D | 92.4% | 90.3% | Qwen3.7 Max |
| HLE | 41.4% | 34.7% | Qwen3.7 Max |
| MMLU-Pro | 89.6% | 88.5% | Qwen3.7 Max |
| MMLU-Redux | 95% | 94.5% | Qwen3.7 Max |
| SuperGPQA | 73.6% | 71.4% | Qwen3.7 Max |
| MMMLU | 90.3% | 89.0% | Qwen3.7 Max |
| GPQA Diamond (Vals) | 90.2% | Coming soon | Coming soon |
| MMLU-Pro (Vals) | 89.3% | Coming soon | Coming soon |

### Instruction Following

- Winner: Tie
- Qwen3.7 Max public-lane score: 91.1 (#16/120)
- Qwen3.7 Plus public-lane score: 91.1 (#17/120)

| Benchmark | Qwen3.7 Max | Qwen3.7 Plus | Winner |
|-----------|-----------|-----------|--------|
| IFEval | 94.3% | 94.6% | Qwen3.7 Plus |
| IFBench | 79.1% | 79.1% | Tie |

### Multilingual

- Winner: Qwen3.7 Max
- Qwen3.7 Max public-lane score: 100 (#1/12)
- Qwen3.7 Plus public-lane score: 78.9 (#3/12)

| Benchmark | Qwen3.7 Max | Qwen3.7 Plus | Winner |
|-----------|-----------|-----------|--------|
| MMLU-ProX | 87% | 85.4% | Qwen3.7 Max |
| NOVA-63 | 59.0% | 58.8% | Qwen3.7 Max |
| INCLUDE | 86.2% | 83.0% | Qwen3.7 Max |
| MAXIFE | 89.2% | 88.8% | Qwen3.7 Max |
| PolyMath | 86.5% | 84.0% | Qwen3.7 Max |

### Mathematics

- Winner: Coming soon
- Qwen3.7 Max public-lane score: 82.1 (Unranked · 3 rankable rows)
- Qwen3.7 Plus public-lane score: 78.3 (Unranked · 3 rankable rows)

| Benchmark | Qwen3.7 Max | Qwen3.7 Plus | Winner |
|-----------|-----------|-----------|--------|
| HMMT Feb 2026 | 97.1% | 92.9% | Qwen3.7 Max |
| IMOAnswerBench | 90.0% | 86.0% | Qwen3.7 Max |
| Apex | 44.5% | 22.7% | Qwen3.7 Max |

## FAQ

### Which is better overall, Qwen3.7 Max or Qwen3.7 Plus?

Qwen3.7 Max is ahead overall on BenchLM right now.

### Where is the biggest gap between Qwen3.7 Max and Qwen3.7 Plus?

The widest category gap is in multilingual, where the averages are 100 for Qwen3.7 Max and 78.9 for Qwen3.7 Plus.

### How many benchmarks does Qwen3.7 Max cover on BenchLM?

Qwen3.7 Max currently has 41 sourced benchmark scores on BenchLM.

### How many benchmarks does Qwen3.7 Plus cover on BenchLM?

Qwen3.7 Plus currently has 52 sourced benchmark scores on BenchLM.

## Related Comparisons

- [Qwen3.7 Max vs Claude Fable 5.1](/compare/claude-fable-5-1-vs-qwen3-7-max)
- [Qwen3.7 Plus vs Claude Fable 5.1](/compare/claude-fable-5-1-vs-qwen3-7-plus)
- [Qwen3.7 Max vs GPT-6 Astra](/compare/gpt-6-astra-vs-qwen3-7-max)
- [Qwen3.7 Plus vs GPT-6 Astra](/compare/gpt-6-astra-vs-qwen3-7-plus)

## Explore More

- [Qwen3.7 Max profile](/models/qwen3-7-max)
- [Qwen3.7 Plus profile](/models/qwen3-7-plus)
- [Compare Pricing](/llm-pricing)
- [Alternative Finder](/tools/alternative-finder)
- [LLM Selector](/tools/llm-selector)
- [Overall Rankings](/best/overall)
