# Atomic Evasion (Atomic evasion)

> Irregular's domain-level evaluation of cybersecurity evasion capability.

Canonical page: https://benchlm.ai/benchmarks/atomicevasion

- Category: [Agentic](/agentic)
- Last updated: September 29, 2026

## About Atomic evasion

- Year: 2026
- Tasks: Atomic cyber tasks
- Format: Domain average
- Difficulty: Cybersecurity evasion
- Paper: [Assessing GPT-5.6 Sol](https://www.irregular.com/research/assessing-gpt-5.6-sol)

This Atomic domain average isolates evasion tasks. BenchLM stores it separately from network and vulnerability-research scores and excludes it from weighted rankings.

Atomic evasion is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Leaderboard (3 models)

| Rank | Model | Creator | Score |
|------|-------|---------|-------|
| 1 | [GPT-5.6 Sol](/models/gpt-5-6-sol) | OpenAI | 56% |
| 2 | [GPT-5.5](/models/gpt-5-5) | OpenAI | 54% |
| 3 | [GPT-6 Astra](/models/gpt-6-astra) | OpenAI | 52% |

## FAQ

### What does Atomic evasion measure?

Irregular's domain-level evaluation of cybersecurity evasion capability.

### Which model scores highest on Atomic evasion?

GPT-5.6 Sol by OpenAI currently leads with a score of 56% on Atomic evasion.

### How many models are evaluated on Atomic evasion?

3 AI models have been evaluated on Atomic evasion on BenchLM.

### Does Atomic evasion affect BenchLM's overall score?

Not directly. Atomic evasion is still displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Compare Top Models on Atomic evasion

- [GPT-5.6 Sol vs GPT-5.5](/compare/gpt-5-5-vs-gpt-5-6-sol)
- [GPT-5.5 vs GPT-6 Astra](/compare/gpt-5-5-vs-gpt-6-astra)
