# SEC-Bench Pro

> Cybersecurity benchmark for agentic vulnerability analysis and exploit-oriented security tasks.

Canonical page: https://benchlm.ai/benchmarks/secbenchpro

- Category: [Agentic](/agentic)
- Last updated: September 27, 2026

## About SEC-Bench Pro

- Year: 2026
- Tasks: Security engineering tasks
- Format: Success rate
- Difficulty: Advanced cybersecurity
- Paper: [Introducing GPT-5.6](https://openai.com/index/gpt-5-6/)

OpenAI reports exact model results on SEC-Bench Pro in its GPT-5.6 launch table. BenchLM stores the provider-run values as display-only cyber evidence until benchmark-native results are available.

SEC-Bench Pro is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Leaderboard (8 models)

| Rank | Model | Creator | Score |
|------|-------|---------|-------|
| 1 | [GPT-6 Astra](/models/gpt-6-astra) | OpenAI | 85.4% |
| 2 | [GPT-5.6 Sol](/models/gpt-5-6-sol) | OpenAI | 71.2% |
| 3 | [GPT-6 Sol](/models/gpt-6-sol) | OpenAI | 66.3% |
| 4 | [MiMo-V2.6-Pro](/models/mimo-v2-6-pro) | Xiaomi | 66.3% |
| 5 | [DeepSeek V4.1 Flash](/models/deepseek-v4-1-flash) | DeepSeek | 62.8% |
| 6 | [MiMo-V2.6-Flash](/models/mimo-v2-6-flash) | Xiaomi | 47.5% |
| 7 | [GPT-5.5](/models/gpt-5-5) | OpenAI | 45.8% |
| 8 | [GPT-6 Luna](/models/gpt-6-luna) | OpenAI | 34.2% |

## FAQ

### What does SEC-Bench Pro measure?

Cybersecurity benchmark for agentic vulnerability analysis and exploit-oriented security tasks.

### Which model scores highest on SEC-Bench Pro?

GPT-6 Astra by OpenAI currently leads with a score of 85.4% on SEC-Bench Pro.

### How many models are evaluated on SEC-Bench Pro?

8 AI models have been evaluated on SEC-Bench Pro on BenchLM.

### Does SEC-Bench Pro affect BenchLM's overall score?

Not directly. SEC-Bench Pro is still displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Compare Top Models on SEC-Bench Pro

- [GPT-6 Astra vs GPT-5.6 Sol](/compare/gpt-5-6-sol-vs-gpt-6-astra)
- [GPT-5.6 Sol vs GPT-6 Sol](/compare/gpt-5-6-sol-vs-gpt-6-sol)
- [GPT-6 Sol vs MiMo-V2.6-Pro](/compare/gpt-6-sol-vs-mimo-v2-6-pro)
- [MiMo-V2.6-Pro vs DeepSeek V4.1 Flash](/compare/deepseek-v4-1-flash-vs-mimo-v2-6-pro)
