Skip to main content

Benchmark profile

Harvard-MIT Mathematics Tournament February 2024 (HMMT Feb 2024)

The 2024 February edition of the Harvard-MIT Mathematics Tournament, continuing the tradition of challenging high school mathematics competition.

Data verified

How BenchLM shows HMMT Feb 2024 right now

BenchLM is tracking HMMT Feb 2024 in the local dataset, but exact-source verification records for these rows are still being attached. To avoid a blank benchmark page, BenchLM shows the current tracked rows below as a display-only reference table.

These tracked rows are useful for inspection and spot-checking, but until exact-source attachments are completed they should not be treated as fully verified public benchmark rows.

105 tracked modelsLocal tracked rowsAwaiting exact-source attachmentsDisplay only

Tracked score on HMMT Feb 2024 — July 20, 2026

BenchLM mirrors the published tracked score view for HMMT Feb 2024. GPT-5.4 leads the public snapshot at 98% , followed by GPT-5.2 Pro (98%) and GPT-5.1-Codex-Max (97%). BenchLM does not use these results to rank models overall.

105 modelsMathRefreshingDisplay onlyUpdated July 20, 2026

Tracked score table (105 models)

Score
1
GPT-5.4OpenAI
98%
2
98%
4
97%
5
97%
6
97%
8
97%
9
GPT-5.1OpenAI
97%
10
GPT-5.2OpenAI
97%
11
97%
12
97%
13
97%
14
97%
15
97%
18
95%
20
93%
21
92%
22
92%
23
91%
25
o3-proOpenAI
88%
26
88%
27
o3OpenAI
86%
28
GLM-5Z.AI
86%
29
86%
31
84%
32
Qwen2.5-1MAlibaba
83%
33
83%
34
82%
35
82%
36
82%
37
81%
38
81%
39
79%
40
Mercury 2Inception
79%
41
78%
42
78%
43
77%
44
Kimi K2.5Moonshot AI
75%
45
74%
46
74%
47
Aion-2.0Aion Labs
72%
48
71%
50
71%
51
Seed 1.6ByteDance
70%
52
Seed-2.0-LiteByteDance
69%
53
68%
55
66%
56
66%
57
66%
59
GPT-4oOpenAI
64%
60
64%
62
63%
63
63%
65
62%
66
62%
68
60%
69
Seed-2.0-MiniByteDance
60%
70
Claude 3 OpusAnthropic
59%
71
58%
72
56%
74
54%
75
52%
76
Moonshot v1Moonshot AI
51%
77
50%
78
49%
79
48%
82
45%
84
LFM2-24B-A2BLiquidAI
44%
85
43%
86
DeepSeek-R1DeepSeek
42%
88
Nova ProAmazon
39%
90
37%
92
35%
93
34%
94
33%
96
31%
98
29%
99
28%
100
28%
101
27%
102
26%
104
22%
105
21%

The published HMMT Feb 2024 snapshot places GPT-5.4 first at 98%. The third row is 1.0 points behind. The broader top-10 range is 1.0 points, so many of the published results sit in a relatively narrow band.

105 models have been evaluated on HMMT Feb 2024. The benchmark falls in the Math category. This category carries a 5% weight in BenchLM.ai's overall scoring system. HMMT Feb 2024 is currently displayed for reference but excluded from the scoring formula, so it does not directly affect overall rankings.

About HMMT Feb 2024

Year

2024

Tasks

Tournament problems

Format

Competition mathematics

Difficulty

High school olympiad level

HMMT Feb 2024 maintains the high standards of mathematical rigor and creativity expected from this premier competition. Problems test advanced mathematical reasoning skills.

BenchLM freshness & provenance

Version

HMMT Feb 2024 2024

Refresh cadence

Annual

Staleness state

Refreshing

Question availability

Public benchmark set

RefreshingDisplay only

BenchLM uses freshness metadata to decide whether a benchmark should still be treated as a strong differentiator, a benchmark to watch, or a display-only reference. For the full scoring policy, see the BenchLM methodology page.

FAQ

What does HMMT Feb 2024 measure?

The 2024 February edition of the Harvard-MIT Mathematics Tournament, continuing the tradition of challenging high school mathematics competition.

Which model leads the published HMMT Feb 2024 snapshot?

GPT-5.4 currently leads the published HMMT Feb 2024 snapshot with 98% tracked score. BenchLM shows this benchmark for display only and does not use it in overall rankings.

How many models are evaluated on HMMT Feb 2024?

105 AI models are included in BenchLM's mirrored HMMT Feb 2024 snapshot, based on the public leaderboard captured on July 20, 2026.

Last updated: July 20, 2026 · mirrored from the public benchmark leaderboard

Choose a model with this week’s evidence

Join 2,000+ readers for ranking moves, pricing changes, and the claims that still need proof.

One email each week. Unsubscribe anytime.