ACE Cyber Range Challenges Solved (ACE solved)
We show this table for reference; we do not rank on it.
Number of advanced cyber-range challenges solved in the joint NIST CAISI and UK AISI preliminary evaluation.
Benchmark score on ACE solved — September 29, 2026
We compile the ACE solved rows from benchmark-owner or independent runs. Kimi K3 leads the table at 0%. We do not use these results to rank models overall.
1 modelAgenticCurrentDisplay onlyUpdated September 29, 2026
Benchmark score table (1 model)
ScoreAbout ACE solved
Year
2026
Tasks
41 advanced cyber-range challenges
Format
Challenges solved
Difficulty
Advanced cyber operations
The preliminary assessment tested 41 ACE challenges. This is a display-only security result from a controlled government evaluation.
Freshness and provenance
Version
ACE solved 2026
Refresh cadence
Quarterly
Staleness state
Current
Question availability
Public benchmark set
BenchLM uses freshness metadata to decide whether a benchmark should still be treated as a strong differentiator, a benchmark to watch, or a display-only reference. For the full scoring policy, see the BenchLM methodology page.
Questions
What does ACE solved measure?
Number of advanced cyber-range challenges solved in the joint NIST CAISI and UK AISI preliminary evaluation.
Which model scores highest on ACE solved?
Kimi K3 by Moonshot AI currently leads with a score of 0% on ACE solved.
How many models are evaluated on ACE solved?
1 AI models have been evaluated on ACE solved on BenchLM.
Know when it’s worth switching models
The model to choose, the cheaper alternative, and the release we would wait on.
Read a sample issueJoin 2,000+ readers.
One email each week. Unsubscribe anytime.