# QwenClawBench

> Qwen's internal OpenClaw-style benchmark for measuring broad real-world agent performance across practical productivity and research tasks.

Canonical page: https://benchlm.ai/benchmarks/qwenclawbench

- Category: [Agentic](/agentic)
- Last updated: September 27, 2026

## About QwenClawBench

- Year: 2026
- Tasks: Real-world agent workflows
- Format: End-to-end agent evaluation
- Difficulty: Broad real-world agentic execution
- Paper: [Qwen3.6 launch benchmarks](https://qwen.ai/blog?id=qwen3.6)

QwenClawBench appears in the Qwen3.6 launch comparisons as an internal real-world agent benchmark. BenchLM tracks it separately rather than merging it into other Claw-style benchmarks because the task mix and exact protocol are Qwen-specific.

QwenClawBench is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Leaderboard (10 models)

| Rank | Model | Creator | Score |
|------|-------|---------|-------|
| 1 | [Qwen3.7 Max](/models/qwen3-7-max) | Alibaba | 64.3% |
| 2 | [Qwen3.7 Plus](/models/qwen3-7-plus) | Alibaba | 61.8% |
| 3 | [Qwen 3.6 Max (preview)](/models/qwen3-6-max-preview) | Alibaba | 59.0% |
| 4 | [Qwen3.6 Plus](/models/qwen3-6-plus) | Alibaba | 57.2% |
| 5 | [Kimi K2.5](/models/kimi-k2-5) | Moonshot AI | 54.3% |
| 6 | [GLM-5](/models/glm-5) | Z.AI | 54.1% |
| 7 | [Qwen3.6-27B](/models/qwen3-6-27b) | Alibaba | 53.4% |
| 8 | [Qwen3.6-35B-A3B](/models/qwen3-6-35b-a3b) | Alibaba | 52.6% |
| 9 | [Claude Opus 4.5](/models/claude-opus-4-5) | Anthropic | 52.3% |
| 10 | [Qwen3.5 397B](/models/qwen3-5-397b) | Alibaba | 51.8% |

## FAQ

### What does QwenClawBench measure?

Qwen's internal OpenClaw-style benchmark for measuring broad real-world agent performance across practical productivity and research tasks.

### Which model scores highest on QwenClawBench?

Qwen3.7 Max by Alibaba currently leads with a score of 64.3% on QwenClawBench.

### How many models are evaluated on QwenClawBench?

10 AI models have been evaluated on QwenClawBench on BenchLM.

### Does QwenClawBench affect BenchLM's overall score?

Not directly. QwenClawBench is still displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Compare Top Models on QwenClawBench

- [Qwen3.7 Max vs Qwen3.7 Plus](/compare/qwen3-7-max-vs-qwen3-7-plus)
- [Qwen3.7 Plus vs Qwen 3.6 Max (preview)](/compare/qwen3-6-max-preview-vs-qwen3-7-plus)
- [Qwen 3.6 Max (preview) vs Qwen3.6 Plus](/compare/qwen3-6-max-preview-vs-qwen3-6-plus)
- [Qwen3.6 Plus vs Kimi K2.5](/compare/kimi-k2-5-vs-qwen3-6-plus)
