# PerceptionBench (Internal) (PerceptionBench)

> Moonshot AI's internal benchmark for atomic visual perception capabilities.

Canonical page: https://benchlm.ai/benchmarks/perceptionbench

- Category: [Multimodal & Grounded](/multimodal-grounded)
- Last updated: September 27, 2026

## About PerceptionBench

- Year: 2026
- Tasks: Internal atomic visual-perception tasks
- Format: Internal evaluation score
- Difficulty: Fine-grained visual perception
- Paper: [Kimi K3: Open Frontier Intelligence](https://www.kimi.com/blog/kimi-k3)

BenchLM stores the provider-published internal benchmark value as display-only evidence and does not treat it as independently reproducible.

PerceptionBench is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Leaderboard (3 models)

| Rank | Model | Creator | Score |
|------|-------|---------|-------|
| 1 | [Qwen3.8 Max](/models/qwen3-8-max) | Alibaba | 63.5% |
| 2 | [Kimi K3](/models/kimi-k3) | Moonshot AI | 58.5% |
| 3 | [dots3-note Preview](/models/dots3-note-preview) | Dots Studio | 53.4% |

## FAQ

### What does PerceptionBench measure?

Moonshot AI's internal benchmark for atomic visual perception capabilities.

### Which model scores highest on PerceptionBench?

Qwen3.8 Max by Alibaba currently leads with a score of 63.5% on PerceptionBench.

### How many models are evaluated on PerceptionBench?

3 AI models have been evaluated on PerceptionBench on BenchLM.

### Does PerceptionBench affect BenchLM's overall score?

Not directly. PerceptionBench is still displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Compare Top Models on PerceptionBench

- [Qwen3.8 Max vs Kimi K3](/compare/kimi-k3-vs-qwen3-8-max)
- [Kimi K3 vs dots3-note Preview](/compare/dots3-note-preview-vs-kimi-k3)
