# Artificial Analysis MATH-500 (AA MATH-500)

> An independently evaluated MATH-500 result from Artificial Analysis.

Canonical page: https://benchlm.ai/benchmarks/aamath500

- Category: [Mathematics](/math)
- Last updated: September 27, 2026

## About AA MATH-500

- Year: 2026
- Tasks: 500 competition mathematics problems
- Format: Accuracy
- Difficulty: High school to undergraduate mathematics
- Paper: [Artificial Analysis MATH-500 Benchmark Leaderboard](https://artificialanalysis.ai/evaluations/math-500)

Stored separately from the core MATH-500 lane because Artificial Analysis controls the evaluation configuration.

AA MATH-500 is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Leaderboard (3 models)

| Rank | Model | Creator | Score |
|------|-------|---------|-------|
| 1 | [GPT-5 (high)](/models/gpt-5-high) | OpenAI | 99.4% |
| 2 | [o3](/models/o3) | OpenAI | 99.2% |
| 3 | [GPT-5 (medium)](/models/gpt-5-medium) | OpenAI | 99.1% |

## FAQ

### What does AA MATH-500 measure?

An independently evaluated MATH-500 result from Artificial Analysis.

### Which model scores highest on AA MATH-500?

GPT-5 (high) by OpenAI currently leads with a score of 99.4% on AA MATH-500.

### How many models are evaluated on AA MATH-500?

3 AI models have been evaluated on AA MATH-500 on BenchLM.

### Does AA MATH-500 affect BenchLM's overall score?

Not directly. AA MATH-500 is still displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Compare Top Models on AA MATH-500

- [GPT-5 (high) vs o3](/compare/gpt-5-high-vs-o3)
- [o3 vs GPT-5 (medium)](/compare/gpt-5-medium-vs-o3)
