# FrontierMath legacy aggregate (FrontierMath (legacy))

> Legacy FrontierMath values retained for historical model pages. This field is not used in current rankings because it can mix prior benchmark versions and slices.

Canonical page: https://benchlm.ai/benchmarks/frontiermath

- Category: [Mathematics](/math)
- Last updated: September 27, 2026

## About FrontierMath (legacy)

- Year: 2024
- Tasks: Historical aggregate
- Format: Open-ended mathematical reasoning with tool access
- Difficulty: Research-level mathematics
- Paper: [FrontierMath: A Benchmark for Evaluating Advanced Mathematical Reasoning in AI](https://epoch.ai/frontiermath)

Use the versioned FrontierMath v2 fields for ranking. They preserve the corrected v2 Tiers 1-3 and Tier 4 private-set scores separately rather than combining them.

FrontierMath (legacy) is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Leaderboard (7 models)

| Rank | Model | Creator | Score |
|------|-------|---------|-------|
| 1 | [GPT-5.6 Sol](/models/gpt-5-6-sol) | OpenAI | 89% |
| 2 | [GPT-5.6 Terra](/models/gpt-5-6-terra) | OpenAI | 84.9% |
| 3 | [GPT-5.6 Luna](/models/gpt-5-6-luna) | OpenAI | 78.6% |
| 4 | [GPT-5.5 Pro](/models/gpt-5-5-pro) | OpenAI | 52.4% |
| 5 | [GPT-5.5](/models/gpt-5-5) | OpenAI | 51.7% |
| 6 | [GPT-5.4 Pro](/models/gpt-5-4-pro) | OpenAI | 50% |
| 7 | [Claude Opus 4.7 (Adaptive)](/models/claude-opus-4-7-adaptive) | Anthropic | 43.8% |

## FAQ

### What does FrontierMath (legacy) measure?

Legacy FrontierMath values retained for historical model pages. This field is not used in current rankings because it can mix prior benchmark versions and slices.

### Which model scores highest on FrontierMath (legacy)?

GPT-5.6 Sol by OpenAI currently leads with a score of 89% on FrontierMath (legacy).

### How many models are evaluated on FrontierMath (legacy)?

7 AI models have been evaluated on FrontierMath (legacy) on BenchLM.

### Does FrontierMath (legacy) affect BenchLM's overall score?

Not directly. FrontierMath (legacy) is still displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Compare Top Models on FrontierMath (legacy)

- [GPT-5.6 Sol vs GPT-5.6 Terra](/compare/gpt-5-6-sol-vs-gpt-5-6-terra)
- [GPT-5.6 Terra vs GPT-5.6 Luna](/compare/gpt-5-6-luna-vs-gpt-5-6-terra)
- [GPT-5.6 Luna vs GPT-5.5 Pro](/compare/gpt-5-5-pro-vs-gpt-5-6-luna)
- [GPT-5.5 Pro vs GPT-5.5](/compare/gpt-5-5-vs-gpt-5-5-pro)
