# HealthBench length-adjusted score (HealthBench (length-adjusted))

> HealthBench score after applying a verbosity penalty to model responses.

Canonical page: https://benchlm.ai/benchmarks/healthbenchlengthadjusted

- Category: [Knowledge](/knowledge)
- Last updated: September 27, 2026

## About HealthBench (length-adjusted)

- Year: 2026
- Tasks: 5,000 multi-turn patient conversations
- Format: Length-adjusted rubric score
- Difficulty: Realistic healthcare conversations
- Paper: [Claude Opus 5 System Card](https://www-cdn.anthropic.com/c5fbac3f0b1280a933ebd26d3cb8bb9f5bdeaf48/Claude%20Opus%205%20System%20Card.pdf)

Section 8.15.1 reports the length-adjusted result separately from the raw score. All Claude models used max effort, no tools, and five trials.

HealthBench (length-adjusted) is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Leaderboard (5 models)

| Rank | Model | Creator | Score |
|------|-------|---------|-------|
| 1 | [Claude Opus 5.5](/models/claude-opus-5-5) | Anthropic | 60.6% |
| 2 | [GPT-6 Astra](/models/gpt-6-astra) | OpenAI | 58.3% |
| 3 | [Claude Opus 5](/models/claude-opus-5) | Anthropic | 57.8% |
| 4 | [GPT-6 Luna](/models/gpt-6-luna) | OpenAI | 54.5% |
| 5 | [GPT-6 Sol](/models/gpt-6-sol) | OpenAI | 53.2% |

## FAQ

### What does HealthBench (length-adjusted) measure?

HealthBench score after applying a verbosity penalty to model responses.

### Which model scores highest on HealthBench (length-adjusted)?

Claude Opus 5.5 by Anthropic currently leads with a score of 60.6% on HealthBench (length-adjusted).

### How many models are evaluated on HealthBench (length-adjusted)?

5 AI models have been evaluated on HealthBench (length-adjusted) on BenchLM.

### Does HealthBench (length-adjusted) affect BenchLM's overall score?

Not directly. HealthBench (length-adjusted) is still displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Compare Top Models on HealthBench (length-adjusted)

- [Claude Opus 5.5 vs GPT-6 Astra](/compare/claude-opus-5-5-vs-gpt-6-astra)
- [GPT-6 Astra vs Claude Opus 5](/compare/claude-opus-5-vs-gpt-6-astra)
- [Claude Opus 5 vs GPT-6 Luna](/compare/claude-opus-5-vs-gpt-6-luna)
- [GPT-6 Luna vs GPT-6 Sol](/compare/gpt-6-luna-vs-gpt-6-sol)
