Skip to main content

Benchmark profile

Harvard-MIT Mathematics Tournament 2025 Aggregate (HMMT 2025)

The most recent February edition of the Harvard-MIT Mathematics Tournament, featuring the latest challenging problems in competitive mathematics.

Data verified

How BenchLM shows HMMT Feb 2025 right now

BenchLM is tracking HMMT Feb 2025 in the local dataset, but exact-source verification records for these rows are still being attached. To avoid a blank benchmark page, BenchLM shows the current tracked rows below as a display-only reference table.

These tracked rows are useful for inspection and spot-checking, but until exact-source attachments are completed they should not be treated as fully verified public benchmark rows.

106 tracked modelsLocal tracked rowsAwaiting exact-source attachmentsDisplay only

Tracked score on HMMT Feb 2025 — July 20, 2026

BenchLM mirrors the published tracked score view for HMMT Feb 2025. GLM-4.7 leads the public snapshot at 97.1% , followed by GPT-5.4 (97%) and GPT-5.2 Pro (97%). BenchLM does not use these results to rank models overall.

106 modelsMathCurrentDisplay onlyUpdated July 20, 2026

Tracked score table (106 models)

Score
1
97.1%
2
GPT-5.4OpenAI
97%
3
97%
5
96%
6
96%
7
96%
9
96%
10
GPT-5.1OpenAI
96%
11
GPT-5.2OpenAI
96%
12
96%
13
96%
14
96%
15
96%
16
96%
17
95.4%
20
94%
22
92%
23
91%
24
90%
26
o3-proOpenAI
87%
27
87%
28
o3OpenAI
85%
29
GLM-5Z.AI
85%
30
85%
32
Qwen2.5-1MAlibaba
82%
33
82%
34
81%
35
81%
36
81%
37
80%
38
80%
39
78%
40
Mercury 2Inception
78%
41
77%
42
77%
43
76%
44
Kimi K2.5Moonshot AI
74%
45
73%
46
73%
47
Aion-2.0Aion Labs
71%
48
70%
49
70%
50
Seed 1.6ByteDance
69%
51
Seed-2.0-LiteByteDance
68%
53
67%
55
65%
56
65%
57
65%
59
GPT-4oOpenAI
63%
60
63%
62
62%
63
62%
65
61%
66
61%
68
59%
69
Seed-2.0-MiniByteDance
59%
70
Claude 3 OpusAnthropic
58%
71
57%
72
55%
74
53%
75
51%
76
Moonshot v1Moonshot AI
50%
77
49%
78
48%
79
47%
82
44%
84
LFM2-24B-A2BLiquidAI
43%
85
42%
86
DeepSeek-R1DeepSeek
41%
88
Kimi K2Moonshot AI
38.8%
89
Nova ProAmazon
38%
91
36%
93
34%
94
33%
95
32%
97
30%
99
28%
100
27%
101
27%
102
26%
103
25%
105
21%
106
20%

The published HMMT Feb 2025 snapshot places GLM-4.7 first at 97.1%. The third row is 0.1 points behind. The broader top-10 range is 1.1 points, so many of the published results sit in a relatively narrow band.

106 models have been evaluated on HMMT Feb 2025. The benchmark falls in the Math category. This category carries a 5% weight in BenchLM.ai's overall scoring system. HMMT Feb 2025 is currently displayed for reference but excluded from the scoring formula, so it does not directly affect overall rankings.

About HMMT Feb 2025

Year

2025

Tasks

Tournament problems

Format

Competition mathematics

Difficulty

High school olympiad level

HMMT Feb 2025 represents the current pinnacle of high school mathematics competition, with problems designed to challenge the brightest mathematical minds.

BenchLM freshness & provenance

Version

HMMT Feb 2025 2025

Refresh cadence

Quarterly

Staleness state

Current

Question availability

Public benchmark set

CurrentDisplay only

BenchLM uses freshness metadata to decide whether a benchmark should still be treated as a strong differentiator, a benchmark to watch, or a display-only reference. For the full scoring policy, see the BenchLM methodology page.

FAQ

What does HMMT Feb 2025 measure?

The most recent February edition of the Harvard-MIT Mathematics Tournament, featuring the latest challenging problems in competitive mathematics.

Which model leads the published HMMT Feb 2025 snapshot?

GLM-4.7 currently leads the published HMMT Feb 2025 snapshot with 97.1% tracked score. BenchLM shows this benchmark for display only and does not use it in overall rankings.

How many models are evaluated on HMMT Feb 2025?

106 AI models are included in BenchLM's mirrored HMMT Feb 2025 snapshot, based on the public leaderboard captured on July 20, 2026.

Last updated: July 20, 2026 · mirrored from the public benchmark leaderboard

Choose a model with this week’s evidence

Join 2,000+ readers for ranking moves, pricing changes, and the claims that still need proof.

One email each week. Unsubscribe anytime.