# Multimodal Multi-disciplinary Video Understanding (MMVU)

> A benchmark for evaluating multimodal models on video understanding tasks across multiple disciplines, emphasizing temporal reasoning and comprehension over video content.

Canonical page: https://benchlm.ai/benchmarks/mmvu

- Category: [Multimodal & Grounded](/multimodal-grounded)
- Last updated: October 7, 2026

## About MMVU

- Year: 2026
- Tasks: Video understanding
- Format: Video reasoning benchmark
- Difficulty: Multi-disciplinary multimodal video reasoning
- Paper: [Kimi K2.5 benchmark release surface](https://www.kimi.com/blog/kimi-k2-5.html)

MMVU is a useful video-understanding benchmark for BenchLM because it appears in frontier provider tables and complements Video-MME with a more discipline-oriented video reasoning slice.

MMVU is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Leaderboard (7 models)

| Rank | Model | Creator | Score |
|------|-------|---------|-------|
| 1 | [Qwen3.8 Max](/models/qwen3-8-max) | Alibaba | 82.4% |
| 2 | [GLM-5.3-Flash](/models/glm-5-3-flash) | Z.AI | 80.5% |
| 3 | [Kimi K2.5](/models/kimi-k2-5) | Moonshot AI | 80.4% |
| 4 | [dots3-note Preview](/models/dots3-note-preview) | Dots Studio | 79.9% |
| 5 | [Qwen3.5-122B-A10B](/models/qwen3-5-122b-a10b) | Alibaba | 74.7% |
| 6 | [Qwen3.5-27B](/models/qwen3-5-27b) | Alibaba | 73.3% |
| 7 | [Qwen3.5-35B-A3B](/models/qwen3-5-35b-a3b) | Alibaba | 72.3% |

## FAQ

### What does MMVU measure?

A benchmark for evaluating multimodal models on video understanding tasks across multiple disciplines, emphasizing temporal reasoning and comprehension over video content.

### Which model scores highest on MMVU?

Qwen3.8 Max by Alibaba currently leads with a score of 82.4% on MMVU.

### How many models are evaluated on MMVU?

7 AI models have published results on MMVU in the BenchLM catalog.

### Does MMVU affect BenchLM's overall score?

Not directly. MMVU is still displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Compare Top Models on MMVU

- [Qwen3.8 Max vs GLM-5.3-Flash](/compare/glm-5-3-flash-vs-qwen3-8-max)
- [GLM-5.3-Flash vs Kimi K2.5](/compare/glm-5-3-flash-vs-kimi-k2-5)
- [Kimi K2.5 vs dots3-note Preview](/compare/dots3-note-preview-vs-kimi-k2-5)
- [dots3-note Preview vs Qwen3.5-122B-A10B](/compare/dots3-note-preview-vs-qwen3-5-122b-a10b)
