# Video-MME

> A comprehensive benchmark for multimodal large language models on video understanding, covering temporal reasoning, perception, and question answering over videos.

Canonical page: https://benchlm.ai/benchmarks/videomme

- Category: [Multimodal & Grounded](/multimodal-grounded)
- Last updated: September 27, 2026

## About Video-MME

- Year: 2024
- Tasks: Video understanding
- Format: Video QA and analysis
- Difficulty: Broad multimodal video reasoning
- Paper: [Video-MME benchmark](https://mme-benchmark.github.io/)

BenchLM tracks the aggregate Video-MME row as a display-oriented video benchmark when providers publish a single overall score rather than separate with-subtitle and without-subtitle splits.

Video-MME is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Leaderboard (3 models)

| Rank | Model | Creator | Score |
|------|-------|---------|-------|
| 1 | [Seed 2.1 Pro](/models/seed-2-1-pro) | ByteDance | 89.2% |
| 2 | [Seed 2.1 Turbo](/models/seed-2-1-turbo) | ByteDance | 89.0% |
| 3 | [Kimi K2.5](/models/kimi-k2-5) | Moonshot AI | 87.4% |

## FAQ

### What does Video-MME measure?

A comprehensive benchmark for multimodal large language models on video understanding, covering temporal reasoning, perception, and question answering over videos.

### Which model scores highest on Video-MME?

Seed 2.1 Pro by ByteDance currently leads with a score of 89.2% on Video-MME.

### How many models are evaluated on Video-MME?

3 AI models have been evaluated on Video-MME on BenchLM.

### Does Video-MME affect BenchLM's overall score?

Not directly. Video-MME is still displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Compare Top Models on Video-MME

- [Seed 2.1 Pro vs Seed 2.1 Turbo](/compare/seed-2-1-pro-vs-seed-2-1-turbo)
- [Seed 2.1 Turbo vs Kimi K2.5](/compare/kimi-k2-5-vs-seed-2-1-turbo)
