# CorpusQA 1M

> A million-token CorpusQA long-context question-answering benchmark reported in DeepSeek-V4 model evaluations.

Canonical page: https://benchlm.ai/benchmarks/corpusqa1m

- Category: [Reasoning](/reasoning)
- Last updated: September 15, 2026

## About CorpusQA 1M

- Year: 2026
- Tasks: Million-token corpus question answering
- Format: Long-context QA accuracy
- Difficulty: Million-token long context
- Paper: [DeepSeek-V4 Technical Report](https://huggingface.co/deepseek-ai/DeepSeek-V4-Pro/blob/main/DeepSeek_V4.pdf)

CorpusQA 1M is tracked as a display-only provider-table row for long-context DeepSeek-V4 comparisons.

CorpusQA 1M is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Leaderboard (2 models)

| Rank | Model | Creator | Score |
|------|-------|---------|-------|
| 1 | [DeepSeek V4 Pro 0813](/models/deepseek-v4-pro-0813) | DeepSeek | 62.0% |
| 2 | [DeepSeek V4 Flash 0731](/models/deepseek-v4-flash-0731) | DeepSeek | 60.5% |

## FAQ

### What does CorpusQA 1M measure?

A million-token CorpusQA long-context question-answering benchmark reported in DeepSeek-V4 model evaluations.

### Which model scores highest on CorpusQA 1M?

DeepSeek V4 Pro 0813 by DeepSeek currently leads with a score of 62.0% on CorpusQA 1M.

### How many models are evaluated on CorpusQA 1M?

2 AI models have been evaluated on CorpusQA 1M on BenchLM.

### Does CorpusQA 1M affect BenchLM's overall score?

Not directly. CorpusQA 1M is still displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Compare Top Models on CorpusQA 1M

- [DeepSeek V4 Pro 0813 vs DeepSeek V4 Flash 0731](/compare/deepseek-v4-flash-0731-vs-deepseek-v4-pro-0813)
