# ArXivMath June 2026 without tools (ArXivMath Jun. 2026 (no tools))

> Final-answer research mathematics problems drawn from recent arXiv abstracts.

Canonical page: https://benchlm.ai/benchmarks/arxivmathjune2026

- Category: [Mathematics](/math)
- Last updated: September 27, 2026

## About ArXivMath Jun. 2026 (no tools)

- Year: 2026
- Tasks: 49 recent research-mathematics problems
- Format: Final-answer accuracy
- Difficulty: Research mathematics
- Paper: [Claude Opus 5 System Card](https://www-cdn.anthropic.com/c5fbac3f0b1280a933ebd26d3cb8bb9f5bdeaf48/Claude%20Opus%205%20System%20Card.pdf)

Section 8.8 reports four-run average accuracy on the 49-problem June 2026 release at max effort without tools.

ArXivMath Jun. 2026 (no tools) is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Leaderboard (1 models)

| Rank | Model | Creator | Score |
|------|-------|---------|-------|
| 1 | [Claude Opus 5](/models/claude-opus-5) | Anthropic | 90.8% |

## FAQ

### What does ArXivMath Jun. 2026 (no tools) measure?

Final-answer research mathematics problems drawn from recent arXiv abstracts.

### Which model scores highest on ArXivMath Jun. 2026 (no tools)?

Claude Opus 5 by Anthropic currently leads with a score of 90.8% on ArXivMath Jun. 2026 (no tools).

### How many models are evaluated on ArXivMath Jun. 2026 (no tools)?

1 AI models have been evaluated on ArXivMath Jun. 2026 (no tools) on BenchLM.

### Does ArXivMath Jun. 2026 (no tools) affect BenchLM's overall score?

Not directly. ArXivMath Jun. 2026 (no tools) is still displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.
