# VideoMMMU

> A video extension of MMMU-style multimodal reasoning over expert questions grounded in temporal media.

Canonical page: https://benchlm.ai/benchmarks/videommmu

- Category: [Multimodal & Grounded](/multimodal-grounded)
- Last updated: September 27, 2026

## About VideoMMMU

- Year: 2026
- Tasks: Video-grounded expert reasoning
- Format: Video + text reasoning
- Difficulty: Frontier multimodal video reasoning
- Paper: [Qwen3.6 launch benchmarks](https://qwen.ai/blog?id=qwen3.6)

VideoMMMU tests whether multimodal reasoning skills extend from static images into temporal video understanding. It is useful for evaluating long-form visual reasoning rather than static scene recognition.

VideoMMMU is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Leaderboard (11 models)

| Rank | Model | Creator | Score |
|------|-------|---------|-------|
| 1 | [Qwen3.8 Max](/models/qwen3-8-max) | Alibaba | 88.7% |
| 2 | [Gemini 3 Pro](/models/gemini-3-pro) | Google | 87.6% |
| 3 | [dots3-note Preview](/models/dots3-note-preview) | Dots Studio | 86.8% |
| 4 | [Kimi K2.5](/models/kimi-k2-5) | Moonshot AI | 86.6% |
| 5 | [Qwen3.7 Plus](/models/qwen3-7-plus) | Alibaba | 85.4% |
| 6 | [Qwen3.5 397B](/models/qwen3-5-397b) | Alibaba | 84.7% |
| 7 | [MiniMax M3](/models/minimax-m3) | MiniMax | 84.6% |
| 8 | [Claude Opus 4.5](/models/claude-opus-4-5) | Anthropic | 84.4% |
| 9 | [Qwen3.6-27B](/models/qwen3-6-27b) | Alibaba | 84.4% |
| 10 | [Qwen3.6 Plus](/models/qwen3-6-plus) | Alibaba | 84.0% |
| 11 | [Qwen3.6-35B-A3B](/models/qwen3-6-35b-a3b) | Alibaba | 83.7% |

## FAQ

### What does VideoMMMU measure?

A video extension of MMMU-style multimodal reasoning over expert questions grounded in temporal media.

### Which model scores highest on VideoMMMU?

Qwen3.8 Max by Alibaba currently leads with a score of 88.7% on VideoMMMU.

### How many models are evaluated on VideoMMMU?

11 AI models have been evaluated on VideoMMMU on BenchLM.

### Does VideoMMMU affect BenchLM's overall score?

Not directly. VideoMMMU is still displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Compare Top Models on VideoMMMU

- [Qwen3.8 Max vs Gemini 3 Pro](/compare/gemini-3-pro-vs-qwen3-8-max)
- [Gemini 3 Pro vs dots3-note Preview](/compare/dots3-note-preview-vs-gemini-3-pro)
- [dots3-note Preview vs Kimi K2.5](/compare/dots3-note-preview-vs-kimi-k2-5)
- [Kimi K2.5 vs Qwen3.7 Plus](/compare/kimi-k2-5-vs-qwen3-7-plus)
