# LVBench

> A long-video understanding benchmark for retrieving and reasoning over information distributed across extended video inputs.

Canonical page: https://benchlm.ai/benchmarks/lvbench

- Category: [Multimodal & Grounded](/multimodal-grounded)
- Last updated: September 27, 2026

## About LVBench

- Year: 2026
- Tasks: Long-form video question answering
- Format: Long-video understanding score
- Difficulty: Extended temporal reasoning
- Paper: [Qwen3.8-Max: A New Bar for Coding and Cowork](https://qwen.ai/blog?id=qwen3.8)

We store Qwen's standard LVBench result without the separate memory system. The provider also reports an LVBench-with-memory row, which remains distinct and is not collapsed into this value.

LVBench is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Leaderboard (5 models)

| Rank | Model | Creator | Score |
|------|-------|---------|-------|
| 1 | [Gemini 3.8 Flash](/models/gemini-3-8-flash) | Google | 87.1% |
| 2 | [Gemini 3.7 Flash](/models/gemini-3-7-flash) | Google | 85.4% |
| 3 | [Qwen3.8 Max](/models/qwen3-8-max) | Alibaba | 81.8% |
| 4 | [Qwen3.8-Omni-Flash](/models/qwen3-8-omni-flash) | Alibaba | 76.9% |
| 5 | [Qwen3.8-Flash-Next](/models/qwen3-8-flash-next) | Alibaba | 76.6% |

## FAQ

### What does LVBench measure?

A long-video understanding benchmark for retrieving and reasoning over information distributed across extended video inputs.

### Which model scores highest on LVBench?

Gemini 3.8 Flash by Google currently leads with a score of 87.1% on LVBench.

### How many models are evaluated on LVBench?

5 AI models have been evaluated on LVBench on BenchLM.

### Does LVBench affect BenchLM's overall score?

Not directly. LVBench is still displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Compare Top Models on LVBench

- [Gemini 3.8 Flash vs Gemini 3.7 Flash](/compare/gemini-3-7-flash-vs-gemini-3-8-flash)
- [Gemini 3.7 Flash vs Qwen3.8 Max](/compare/gemini-3-7-flash-vs-qwen3-8-max)
- [Qwen3.8 Max vs Qwen3.8-Omni-Flash](/compare/qwen3-8-max-vs-qwen3-8-omni-flash)
- [Qwen3.8-Omni-Flash vs Qwen3.8-Flash-Next](/compare/qwen3-8-flash-next-vs-qwen3-8-omni-flash)
