# SimpleVQA

> A visual question answering benchmark focused on straightforward image-grounded understanding.

Canonical page: https://benchlm.ai/benchmarks/simplevqa

- Category: [Multimodal & Grounded](/multimodal-grounded)
- Last updated: September 27, 2026

## About SimpleVQA

- Year: 2026
- Tasks: Visual QA tasks
- Format: Image-grounded question answering
- Difficulty: General visual understanding
- Paper: [GLM-5V-Turbo](https://docs.z.ai/guides/vlm/glm-5v-turbo)

BenchLM uses SimpleVQA as a display-only visual QA reference rather than a weighted multimodal ranking input.

SimpleVQA is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Leaderboard (11 models)

| Rank | Model | Creator | Score |
|------|-------|---------|-------|
| 1 | [Qwen3.7 Plus](/models/qwen3-7-plus) | Alibaba | 81.7% |
| 2 | [Step 3.7 Flash](/models/step-3-7-flash) | StepFun | 79.2% |
| 3 | [Qwen3.8 Max](/models/qwen3-8-max) | Alibaba | 75.0% |
| 4 | [dots3-note Preview](/models/dots3-note-preview) | Dots Studio | 72.5% |
| 5 | [Gemini 3.1 Pro](/models/gemini-3-1-pro) | Google | 72.4% |
| 6 | [Muse Spark](/models/muse-spark) | Meta | 71.3% |
| 7 | [GPT-5.4](/models/gpt-5-4) | OpenAI | 61.1% |
| 8 | [Qwen3.6-35B-A3B](/models/qwen3-6-35b-a3b) | Alibaba | 58.9% |
| 9 | [Grok 4.20](/models/grok-4-20-beta) | xAI | 57.4% |
| 10 | [Qwen3.6-27B](/models/qwen3-6-27b) | Alibaba | 56.1% |
| 11 | [LFM2.5-VL-3B](/models/lfm2-5-vl-3b) | LiquidAI | 35.4% |

## FAQ

### What does SimpleVQA measure?

A visual question answering benchmark focused on straightforward image-grounded understanding.

### Which model scores highest on SimpleVQA?

Qwen3.7 Plus by Alibaba currently leads with a score of 81.7% on SimpleVQA.

### How many models are evaluated on SimpleVQA?

11 AI models have been evaluated on SimpleVQA on BenchLM.

### Does SimpleVQA affect BenchLM's overall score?

Not directly. SimpleVQA is still displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Compare Top Models on SimpleVQA

- [Qwen3.7 Plus vs Step 3.7 Flash](/compare/qwen3-7-plus-vs-step-3-7-flash)
- [Step 3.7 Flash vs Qwen3.8 Max](/compare/qwen3-8-max-vs-step-3-7-flash)
- [Qwen3.8 Max vs dots3-note Preview](/compare/dots3-note-preview-vs-qwen3-8-max)
- [dots3-note Preview vs Gemini 3.1 Pro](/compare/dots3-note-preview-vs-gemini-3-1-pro)
