# MM-ClawBench

> An OpenClaw-derived agent benchmark covering practical work and life tasks such as office document delivery, research, planning, and code maintenance.

Canonical page: https://benchlm.ai/benchmarks/mmclawbench

- Category: [Agentic](/agentic)
- Last updated: September 27, 2026

## About MM-ClawBench

- Year: 2026
- Tasks: OpenClaw-style real-world tasks
- Format: Agent workflow evaluation
- Difficulty: Broad real-world agentic execution
- Paper: [MiniMax M2.7: Early Echoes of Self-Evolution](https://www.minimax.io/news/minimax-m27-en)

MiniMax built MM-ClawBench from commonly used OpenClaw tasks to evaluate how well models handle broad real-world agent scenarios across work and personal productivity.

MM-ClawBench is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Leaderboard (2 models)

| Rank | Model | Creator | Score |
|------|-------|---------|-------|
| 1 | [MiniMax M2.7](/models/minimax-m2-7) | MiniMax | 62.7% |
| 2 | [MiMo-V2.5](/models/mimo-v2-5) | Xiaomi | 23.8% |

## FAQ

### What does MM-ClawBench measure?

An OpenClaw-derived agent benchmark covering practical work and life tasks such as office document delivery, research, planning, and code maintenance.

### Which model scores highest on MM-ClawBench?

MiniMax M2.7 by MiniMax currently leads with a score of 62.7% on MM-ClawBench.

### How many models are evaluated on MM-ClawBench?

2 AI models have been evaluated on MM-ClawBench on BenchLM.

### Does MM-ClawBench affect BenchLM's overall score?

Not directly. MM-ClawBench is still displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Compare Top Models on MM-ClawBench

- [MiniMax M2.7 vs MiMo-V2.5](/compare/mimo-v2-5-vs-minimax-m2-7)
