Harvard-MIT Mathematics Tournament February 2026 (HMMT Feb 2026)
A February 2026 HMMT slice used in newer frontier-model math comparisons.
Data verified 33 confirmed releases in the last 30 daysSee provider release alertsTop models on HMMT Feb 2026 — September 15, 2026
As of September 15, 2026, Qwen3.7 Max leads the HMMT Feb 2026 leaderboard with 97.1% , followed by DeepSeek V4 Pro 0813 (95.2%) and DeepSeek V4 Flash 0731 (94.8%).
Qwen3.7 Max
Alibaba
DeepSeek V4 Pro 0813
DeepSeek
DeepSeek V4 Flash 0731
DeepSeek
27 modelsMath25% of category scoreCurrentUpdated September 15, 2026
Leaderboard (27 models)
ScoreAccording to BenchLM.ai, Qwen3.7 Max leads the HMMT Feb 2026 benchmark with a score of 97.1%, followed by DeepSeek V4 Pro 0813 (95.2%) and DeepSeek V4 Flash 0731 (94.8%). The top models are clustered within 2.3 points, suggesting this benchmark is nearing saturation for frontier models.
27 models have been evaluated on HMMT Feb 2026. The benchmark falls in the Math category. This category carries a 5% weight in BenchLM.ai's overall scoring system. Within that category, HMMT Feb 2026 contributes 25% of the category score, so strong performance here directly affects a model's overall ranking.
About HMMT Feb 2026
Year
2026
Tasks
Competition math problems
Format
Contest mathematics
Difficulty
Olympiad-style mathematics
HMMT February 2026 matters because small score deltas at the frontier often depend on which contest set is used. BenchLM keeps this newer slice distinct from older HMMT summary rows.
BenchLM freshness & provenance
Version
HMMT Feb 2026 2026
Refresh cadence
Quarterly
Staleness state
Current
Question availability
Public benchmark set
BenchLM uses freshness metadata to decide whether a benchmark should still be treated as a strong differentiator, a benchmark to watch, or a display-only reference. For the full scoring policy, see the BenchLM methodology page.
FAQ
What does HMMT Feb 2026 measure?
A February 2026 HMMT slice used in newer frontier-model math comparisons.
Which model scores highest on HMMT Feb 2026?
Qwen3.7 Max by Alibaba currently leads with a score of 97.1% on HMMT Feb 2026.
How many models are evaluated on HMMT Feb 2026?
27 AI models have been evaluated on HMMT Feb 2026 on BenchLM.
Compare Top Models on HMMT Feb 2026
Know when it’s worth switching models
The model to choose, the cheaper alternative, and the release we would wait on.
Read a sample issueJoin 2,000+ readers.
One email each week. Unsubscribe anytime.