Skip to main content

Benchmark profile

Harvard-MIT Mathematics Tournament February 2023 (HMMT Feb 2023)

A prestigious high school mathematics competition hosted jointly by Harvard and MIT, featuring challenging problems across various mathematical disciplines.

Data verified

How BenchLM shows HMMT Feb 2023 right now

BenchLM is tracking HMMT Feb 2023 in the local dataset, but exact-source verification records for these rows are still being attached. To avoid a blank benchmark page, BenchLM shows the current tracked rows below as a display-only reference table.

These tracked rows are useful for inspection and spot-checking, but until exact-source attachments are completed they should not be treated as fully verified public benchmark rows.

105 tracked modelsLocal tracked rowsAwaiting exact-source attachmentsDisplay only

Tracked score on HMMT Feb 2023 — July 20, 2026

BenchLM mirrors the published tracked score view for HMMT Feb 2023. GPT-5.4 leads the public snapshot at 96% , followed by GPT-5.2 Pro (96%) and GPT-5.1-Codex-Max (95%). BenchLM does not use these results to rank models overall.

105 modelsMathStaleDisplay onlyUpdated July 20, 2026

Tracked score table (105 models)

Score
1
GPT-5.4OpenAI
96%
2
96%
4
95%
5
95%
6
95%
8
95%
9
GPT-5.1OpenAI
95%
10
GPT-5.2OpenAI
95%
11
95%
12
95%
13
95%
14
95%
15
95%
18
93%
20
91%
21
90%
22
90%
23
89%
25
o3-proOpenAI
86%
26
86%
27
o3OpenAI
84%
28
GLM-5Z.AI
84%
29
84%
31
82%
32
Qwen2.5-1MAlibaba
81%
33
81%
34
80%
35
80%
36
80%
37
79%
38
79%
39
77%
40
Mercury 2Inception
77%
41
76%
42
76%
43
75%
44
Kimi K2.5Moonshot AI
73%
45
72%
46
72%
47
Aion-2.0Aion Labs
70%
48
69%
50
69%
51
Seed 1.6ByteDance
68%
52
Seed-2.0-LiteByteDance
67%
53
66%
55
64%
56
64%
57
64%
59
GPT-4oOpenAI
62%
60
62%
62
61%
63
61%
65
60%
66
60%
68
58%
69
Seed-2.0-MiniByteDance
58%
70
Claude 3 OpusAnthropic
57%
71
56%
72
54%
74
52%
75
50%
76
Moonshot v1Moonshot AI
49%
77
48%
78
47%
79
46%
82
43%
84
LFM2-24B-A2BLiquidAI
42%
85
41%
86
DeepSeek-R1DeepSeek
40%
88
Nova ProAmazon
37%
90
35%
92
33%
93
32%
94
31%
96
29%
98
27%
99
26%
100
26%
101
25%
102
24%
104
20%
105
19%

The published HMMT Feb 2023 snapshot places GPT-5.4 first at 96%. The third row is 1.0 points behind. The broader top-10 range is 1.0 points, so many of the published results sit in a relatively narrow band.

105 models have been evaluated on HMMT Feb 2023. The benchmark falls in the Math category. This category carries a 5% weight in BenchLM.ai's overall scoring system. HMMT Feb 2023 is currently displayed for reference but excluded from the scoring formula, so it does not directly affect overall rankings.

About HMMT Feb 2023

Year

2023

Tasks

Tournament problems

Format

Competition mathematics

Difficulty

High school olympiad level

HMMT is one of the most competitive high school mathematics tournaments in the US. Problems span algebra, geometry, combinatorics, and number theory, requiring deep mathematical insight.

BenchLM freshness & provenance

Version

HMMT Feb 2023 2023

Refresh cadence

Static

Staleness state

Stale

Question availability

Public benchmark set

StaleDisplay only

BenchLM uses freshness metadata to decide whether a benchmark should still be treated as a strong differentiator, a benchmark to watch, or a display-only reference. For the full scoring policy, see the BenchLM methodology page.

FAQ

What does HMMT Feb 2023 measure?

A prestigious high school mathematics competition hosted jointly by Harvard and MIT, featuring challenging problems across various mathematical disciplines.

Which model leads the published HMMT Feb 2023 snapshot?

GPT-5.4 currently leads the published HMMT Feb 2023 snapshot with 96% tracked score. BenchLM shows this benchmark for display only and does not use it in overall rankings.

How many models are evaluated on HMMT Feb 2023?

105 AI models are included in BenchLM's mirrored HMMT Feb 2023 snapshot, based on the public leaderboard captured on July 20, 2026.

Last updated: July 20, 2026 · mirrored from the public benchmark leaderboard

Choose a model with this week’s evidence

Join 2,000+ readers for ranking moves, pricing changes, and the claims that still need proof.

One email each week. Unsubscribe anytime.