# Apex Shortlist

> A shortlist subset of the Apex mathematical reasoning benchmark reported in DeepSeek-V4 model evaluations.

Canonical page: https://benchlm.ai/benchmarks/apexshortlist

- Category: [Mathematics](/math)
- Last updated: September 15, 2026

## About Apex Shortlist

- Year: 2026
- Tasks: Advanced mathematical reasoning
- Format: Pass@1 math benchmark
- Difficulty: Frontier math reasoning
- Paper: [DeepSeek-V4 Technical Report](https://huggingface.co/deepseek-ai/DeepSeek-V4-Pro/blob/main/DeepSeek_V4.pdf)

BenchLM stores Apex Shortlist separately from the broader Apex row so provider-reported table values remain traceable.

Apex Shortlist is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Leaderboard (3 models)

| Rank | Model | Creator | Score |
|------|-------|---------|-------|
| 1 | [DeepSeek V4 Pro 0813](/models/deepseek-v4-pro-0813) | DeepSeek | 90.2% |
| 2 | [A.X K2](/models/a-x-k2) | SK Telecom | 88.6% |
| 3 | [DeepSeek V4 Flash 0731](/models/deepseek-v4-flash-0731) | DeepSeek | 85.7% |

## FAQ

### What does Apex Shortlist measure?

A shortlist subset of the Apex mathematical reasoning benchmark reported in DeepSeek-V4 model evaluations.

### Which model scores highest on Apex Shortlist?

DeepSeek V4 Pro 0813 by DeepSeek currently leads with a score of 90.2% on Apex Shortlist.

### How many models are evaluated on Apex Shortlist?

3 AI models have been evaluated on Apex Shortlist on BenchLM.

### Does Apex Shortlist affect BenchLM's overall score?

Not directly. Apex Shortlist is still displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Compare Top Models on Apex Shortlist

- [DeepSeek V4 Pro 0813 vs A.X K2](/compare/a-x-k2-vs-deepseek-v4-pro-0813)
- [A.X K2 vs DeepSeek V4 Flash 0731](/compare/a-x-k2-vs-deepseek-v4-flash-0731)
