# CRUXEval (Fastino report)

> Fastino reports GLiDE at 92.6% on CRUXEval in its Decision Index 0.2.1 run. The comparison quotes the published Jev 1.13.0 reference.

Canonical page: https://benchlm.ai/benchmarks/cruxeval-decision-index-fastino

- Category: [Decision Models](/decision-models)
- Last updated: September 30, 2026

## About CRUXEval (Fastino report)

- Year: 2026
- Tasks: Decision Index task conversion
- Format: Raw accuracy (%)
- Difficulty: Varies by task and supplied answer options
- Paper: [Fastino GLiDE launch and Decision Index results](https://fastino.ai/blog/introducing-glide-the-first-thinking-decision-model)

The five area scores and overall index are chance-adjusted skill points. CLadder and CRUXEval are raw accuracies. Fastino reports a complete run, but publishes exact GLiDE values for only these eight metrics. Radar-chart gaps do not supply the other individual benchmark accuracies. We have not rerun the evaluation, and these results stay outside general model rankings. The area chart states RTX PRO 6000 hardware and 155,390 requests for the full suite; that count is not a sample size for every metric. The benchmark-owner artifact confirms the Jev reference identity.

CRUXEval (Fastino report) is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Published results and quoted references

The five area scores and overall index are chance-adjusted skill points. CLadder and CRUXEval are raw accuracies. Fastino reports a complete run, but publishes exact GLiDE values for only these eight metrics. Radar-chart gaps do not supply the other individual benchmark accuracies. We have not rerun the evaluation, and these results stay outside general model rankings. The launch chart also quotes five other overall references, including Liquid AI’s self-reported d1 reproduction; their category and task scores are not inferred.

| Evaluated system | Published score | Source qualification |
|---|---|---|
| [GLiDE](/models/fastino-glide) | 92.6% | Fastino-reported GLiDE run; RTX PRO 6000; official 0.2.1 scorer |
| [TypeSafe Jev 1.13.0](/models/jev-1-13-0) | 73.0% | Published Jev 1.13.0 reference; values quoted by Fastino |

## FAQ

### Is GLiDE on the public Decision Index leaderboard?

Fastino says GLiDE is not listed on the public Decision Index 0.2.1 leaderboard because submissions are paused. These values come from Fastino’s own run with the official scorer, published September 30, 2026. They are a provider report, not an independently reproduced leaderboard entry or a general model ranking.

### Are skill points and benchmark accuracy interchangeable?

Decision Index skill points adjust results for the chance baseline before combining benchmarks. Raw accuracy is the percentage of correct answers under a particular task setup. The overall index and five areas use skill points; the published CLadder and CRUXEval figures use accuracy. Compare each metric only with the same protocol.
