# CTI-REALM

> A cybersecurity benchmark that measures whether an agent can turn raw threat-intelligence reports into working detection rules.

Canonical page: https://benchlm.ai/benchmarks/ctirealm

- Category: [Agentic](/agentic)
- Last updated: September 27, 2026

## About CTI-REALM

- Year: 2026
- Tasks: Threat-intelligence-to-detection-rule workflows
- Format: Success rate
- Difficulty: Professional cyber threat detection
- Paper: [Introducing Fugu-Cyber](https://sakana.ai/fugu-cyber-release/)

Sakana AI describes CTI-REALM as a real-world security benchmark for translating cyber threat intelligence into executable detection rules. We store exact provider-reported results as display-only evidence until the benchmark owner publishes a stable, independently verifiable leaderboard artifact.

CTI-REALM is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Leaderboard (1 models)

| Rank | Model | Creator | Score |
|------|-------|---------|-------|
| 1 | [Fugu Cyber](/models/sakana-fugu-cyber) | Sakana AI | 72.1% |

## FAQ

### What does CTI-REALM measure?

A cybersecurity benchmark that measures whether an agent can turn raw threat-intelligence reports into working detection rules.

### Which model scores highest on CTI-REALM?

Fugu Cyber by Sakana AI currently leads with a score of 72.1% on CTI-REALM.

### How many models are evaluated on CTI-REALM?

1 AI models have been evaluated on CTI-REALM on BenchLM.

### Does CTI-REALM affect BenchLM's overall score?

Not directly. CTI-REALM is still displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.
