Anthropic Protein Design evaluation (Protein Design)
We mirror this table; we do not rank on it.
Generates novel protein sequences under family, topology, globularity, and structural-motif constraints.
Benchmark score on Protein Design — September 22, 2026
We mirror the published score view for Protein Design. Claude Opus 5.5 leads the public snapshot at 60.2%, followed by Claude Opus 5 (42.5%). We do not use these results to rank models overall.
Claude Opus 5.5
Anthropic
Claude Opus 5
Anthropic
2 modelsKnowledgeCurrentDisplay onlyUpdated September 22, 2026
Benchmark score table (2 models)
ScoreAbout Protein Design
Year
2026
Tasks
Constrained protein-sequence design tasks
Format
Composite score
Difficulty
Computational protein design
Section 8.17.4 reports a composite of constraint satisfaction, folding confidence, and sequence novelty without tool access.
Freshness and provenance
Version
Protein Design 2026
Refresh cadence
Quarterly
Staleness state
Current
Question availability
Public benchmark set
BenchLM uses freshness metadata to decide whether a benchmark should still be treated as a strong differentiator, a benchmark to watch, or a display-only reference. For the full scoring policy, see the BenchLM methodology page.
Questions
What does Protein Design measure?
Generates novel protein sequences under family, topology, globularity, and structural-motif constraints.
Which model scores highest on Protein Design?
Claude Opus 5.5 by Anthropic currently leads with a score of 60.2% on Protein Design.
How many models are evaluated on Protein Design?
2 AI models have been evaluated on Protein Design on BenchLM.
Compare top models on Protein Design
Know when it’s worth switching models
The model to choose, the cheaper alternative, and the release we would wait on.
Read a sample issueJoin 2,000+ readers.
One email each week. Unsubscribe anytime.