Skip to main content
Radar

Keep up with the models you depend on. Follow price changes, retirements, and API updates.Follow the models you depend on.

Follow model changes

Codeforces Rating (Codeforces)

Competitive-programming rating reported for DeepSeek-V4 thinking-mode evaluations.

Data verified 37 confirmed releases in the last 30 daysSee provider release alerts

Benchmark score on Codeforces — September 10, 2026

We mirror the published score view for Codeforces. DeepSeek V4.1 Flash leads the public snapshot at 3471.0, followed by DeepSeek V4 Pro 0813 (3206.0) and dots3-note Preview (3056.0). We do not use these results to rank models overall.

6 modelsCodingCurrentDisplay onlyUpdated September 10, 2026

Benchmark score table (6 models)

Score
1
DeepSeek V4.1 FlashDeepSeek · Open weight
3471.0
2
DeepSeek V4 Pro 0813DeepSeek · Closed
3206.0
3
dots3-note PreviewDots Studio · Open weight
3056.0
4
DeepSeek V4 Flash 0731DeepSeek · Closed
3052.0
5
DeepSeek V4 Pro (High)DeepSeek · Open weight
2919.0
6
DeepSeek V4 Flash (High)DeepSeek · Closed
2816.0

The published Codeforces snapshot places DeepSeek V4.1 Flash first at 3471.0. The third row is 415.0 score units behind. The broader top-10 range is 655.0 score units, so the table still separates the published systems.

6 models have been evaluated on Codeforces. The benchmark falls in the Coding category. This category carries a 20% weight in BenchLM.ai's overall scoring system. Codeforces is currently displayed for reference but excluded from the scoring formula, so it does not directly affect overall rankings.

About Codeforces

Year

2026

Tasks

Competitive programming contests

Format

Rating

Difficulty

Elite competitive programming

BenchLM stores Codeforces as a display-only provider-table row because its rating scale is not a 0-100 percentage benchmark.

BenchLM freshness & provenance

Version

Codeforces 2026

Refresh cadence

Quarterly

Staleness state

Current

Question availability

Public benchmark set

CurrentDisplay only

BenchLM uses freshness metadata to decide whether a benchmark should still be treated as a strong differentiator, a benchmark to watch, or a display-only reference. For the full scoring policy, see the BenchLM methodology page.

FAQ

What does Codeforces measure?

Competitive-programming rating reported for DeepSeek-V4 thinking-mode evaluations.

Which model scores highest on Codeforces?

DeepSeek V4.1 Flash by DeepSeek currently leads with a score of 3471.0 on Codeforces.

How many models are evaluated on Codeforces?

6 AI models have been evaluated on Codeforces on BenchLM.

Last updated: September 10, 2026 · BenchLM version Codeforces 2026

Know when it’s worth switching models

The model to choose, the cheaper alternative, and the release we would wait on.

Read a sample issue

Join 2,000+ readers.

One email each week. Unsubscribe anytime.