HAE-RAE Math 8K (HRM8K)
We show this table for reference; we do not rank on it.
Korean mathematical reasoning (high-school to Olympiad level).
Benchmark score on HRM8K — September 29, 2026
We compile the HRM8K rows from provider self-reports. Solar Open 2 leads the table at 92.2%. We do not use these results to rank models overall.
1 modelKoreanKorean-language benchmarkRefreshingDisplay onlyUpdated September 29, 2026
Benchmark score table (1 model)
ScoreAbout HRM8K
Tasks
8,011 instances
Format
Math word problems
Difficulty
Olympiad level
Freshness and provenance
Version
HRM8K
Refresh cadence
Static
Staleness state
Refreshing
Question availability
Public benchmark set
BenchLM uses freshness metadata to decide whether a benchmark should still be treated as a strong differentiator, a benchmark to watch, or a display-only reference. For the full scoring policy, see the BenchLM methodology page.
Questions
What does HRM8K measure?
Korean mathematical reasoning (high-school to Olympiad level).
Which model scores highest on HRM8K?
Solar Open 2 by Upstage currently leads with a score of 92.2% on HRM8K.
How many models are evaluated on HRM8K?
1 AI models have been evaluated on HRM8K on BenchLM.
Know when it’s worth switching models
The model to choose, the cheaper alternative, and the release we would wait on.
Read a sample issueJoin 2,000+ readers.
One email each week. Unsubscribe anytime.