# WinoGrande

> A commonsense coreference benchmark reported in DeepSeek-V4 base-model evaluations.

Canonical page: https://benchlm.ai/benchmarks/winogrande

- Category: [Reasoning](/reasoning)
- Last updated: September 15, 2026

## About WinoGrande

- Year: 2026
- Tasks: Coreference resolution questions
- Format: Exact match
- Difficulty: Commonsense reasoning
- Paper: [DeepSeek-V4 Technical Report](https://huggingface.co/deepseek-ai/DeepSeek-V4-Pro/blob/main/DeepSeek_V4.pdf)

BenchLM stores WinoGrande as a display-only provider-table row when exact values are published in DeepSeek-V4 evaluations.

WinoGrande is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Leaderboard (0 models)

Benchmark data for this page is coming soon.

## FAQ

### What does WinoGrande measure?

A commonsense coreference benchmark reported in DeepSeek-V4 base-model evaluations.

### Which model scores highest on WinoGrande?

No models have been evaluated on WinoGrande yet.

### How many models are evaluated on WinoGrande?

0 AI models have been evaluated on WinoGrande on BenchLM.

### Does WinoGrande affect BenchLM's overall score?

Not directly. WinoGrande is still displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.
