# DeepSeek DSBench Hard (DSBench-Hard)

> DeepSeek's internal hard coding-agent benchmark.

Canonical page: https://benchlm.ai/benchmarks/dsbenchhard

- Category: [Coding](/coding)
- Last updated: September 27, 2026

## About DSBench-Hard

- Year: 2026
- Tasks: Internal hard coding-agent tasks
- Format: Provider-reported score
- Difficulty: Advanced coding-agent challenges
- Paper: [DeepSeek V4 Flash 0731 update](https://api-docs.deepseek.com/zh-cn/updates/)

DeepSeek labels DSBench-Hard as an internal benchmark. BenchLM stores the provider-reported exact result as display-only launch evidence and does not compare it with public benchmark leaderboards.

DSBench-Hard is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Leaderboard (2 models)

| Rank | Model | Creator | Score |
|------|-------|---------|-------|
| 1 | [DeepSeek V4 Pro 0813](/models/deepseek-v4-pro-0813) | DeepSeek | 67.2% |
| 2 | [DeepSeek V4 Flash 0731](/models/deepseek-v4-flash-0731) | DeepSeek | 59.6% |

## FAQ

### What does DSBench-Hard measure?

DeepSeek's internal hard coding-agent benchmark.

### Which model scores highest on DSBench-Hard?

DeepSeek V4 Pro 0813 by DeepSeek currently leads with a score of 67.2% on DSBench-Hard.

### How many models are evaluated on DSBench-Hard?

2 AI models have been evaluated on DSBench-Hard on BenchLM.

### Does DSBench-Hard affect BenchLM's overall score?

Not directly. DSBench-Hard is still displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Compare Top Models on DSBench-Hard

- [DeepSeek V4 Pro 0813 vs DeepSeek V4 Flash 0731](/compare/deepseek-v4-flash-0731-vs-deepseek-v4-pro-0813)
