# Vision2Web

> A benchmark for converting visual references into functional web implementations.

Canonical page: https://benchlm.ai/benchmarks/vision2web

- Category: [Multimodal & Grounded](/multimodal-grounded)
- Last updated: September 27, 2026

## About Vision2Web

- Year: 2026
- Tasks: Screenshot-to-web tasks
- Format: Visual reference to web implementation
- Difficulty: Multimodal web generation
- Paper: [GLM-5V-Turbo](https://docs.z.ai/guides/vlm/glm-5v-turbo)

BenchLM stores Vision2Web as a display-only screenshot-to-web benchmark reference outside the weighted core schema.

Vision2Web is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Leaderboard (4 models)

| Rank | Model | Creator | Score |
|------|-------|---------|-------|
| 1 | [Qwen3.8 Max](/models/qwen3-8-max) | Alibaba | 69.0% |
| 2 | [Qwen3.8-Flash-Next](/models/qwen3-8-flash-next) | Alibaba | 64.0% |
| 3 | [Qwen3.8-27B](/models/qwen3-8-27b) | Alibaba | 62.9% |
| 4 | [Qwen3.8-Omni-Flash](/models/qwen3-8-omni-flash) | Alibaba | 62.9% |

## FAQ

### What does Vision2Web measure?

A benchmark for converting visual references into functional web implementations.

### Which model scores highest on Vision2Web?

Qwen3.8 Max by Alibaba currently leads with a score of 69.0% on Vision2Web.

### How many models are evaluated on Vision2Web?

4 AI models have been evaluated on Vision2Web on BenchLM.

### Does Vision2Web affect BenchLM's overall score?

Not directly. Vision2Web is still displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Compare Top Models on Vision2Web

- [Qwen3.8 Max vs Qwen3.8-Flash-Next](/compare/qwen3-8-flash-next-vs-qwen3-8-max)
- [Qwen3.8-Flash-Next vs Qwen3.8-27B](/compare/qwen3-8-27b-vs-qwen3-8-flash-next)
- [Qwen3.8-27B vs Qwen3.8-Omni-Flash](/compare/qwen3-8-27b-vs-qwen3-8-omni-flash)
