Atomic Evasion (Atomic evasion)
Irregular's domain-level evaluation of cybersecurity evasion capability.
Benchmark score on Atomic evasion — August 13, 2026
BenchLM mirrors the published score view for Atomic evasion. GPT-5.6 Sol leads the public snapshot at 56% , followed by GPT-5.5 (54%). We do not use these results to rank models overall.
GPT-5.6 Sol
OpenAI
gpt-5-6-sol
GPT-5.5
OpenAI
gpt-5-5
Benchmark score table (2 models)
ScoreAbout Atomic evasion
Year
2026
Tasks
Atomic cyber tasks
Format
Domain average
Difficulty
Cybersecurity evasion
This Atomic domain average isolates evasion tasks. BenchLM stores it separately from network and vulnerability-research scores and excludes it from weighted rankings.
BenchLM freshness & provenance
Version
Atomic evasion 2026
Refresh cadence
Quarterly
Staleness state
Current
Question availability
Public benchmark set
BenchLM uses freshness metadata to decide whether a benchmark should still be treated as a strong differentiator, a benchmark to watch, or a display-only reference. For the full scoring policy, see the BenchLM methodology page.
FAQ
What does Atomic evasion measure?
Irregular's domain-level evaluation of cybersecurity evasion capability.
Which model scores highest on Atomic evasion?
GPT-5.6 Sol by OpenAI currently leads with a score of 56% on Atomic evasion.
How many models are evaluated on Atomic evasion?
2 AI models have been evaluated on Atomic evasion on BenchLM.
Compare Top Models on Atomic evasion
Know when it’s worth switching models
The model to choose, the cheaper alternative, and the release we would wait on.
Read a sample issueJoin 2,000+ readers.
One email each week. Unsubscribe anytime.