# GPT-6.1 Sol Benchmark Scores & Performance

> GPT-6.1 Sol by OpenAI scores 66.86/100 overall, ranking #23 out of 514 AI models.

Canonical page: https://benchlm.ai/models/gpt-6-1-sol

Last updated: September 29, 2026

## Model Details

| Property | Value |
|----------|-------|
| Creator | OpenAI |
| Source Type | Proprietary |
| Reasoning Type | Reasoning |
| Context Window | 1.05M |
| Official model card | [OpenAI GPT-6.1 Sol model documentation](https://developers.openai.com/api/docs/models/gpt-6.1-sol) |
| Overall Score | 66.86/100 |
| Overall Rank | #23 of 514 |

## Family & Coverage

- Family: GPT-6
- Variant: sol-6-1 (sol-6-1)
- Benchmarks covered: 21 of 486
- Sibling models: [GPT-6 Astra](/models/gpt-6-astra), [GPT-6 Sol](/models/gpt-6-sol), [GPT-6 Luna](/models/gpt-6-luna)
- Related earlier model: [GPT-6 Sol](/models/gpt-6-sol)
- Coverage note: BenchLM currently has partial benchmark coverage for this model, so the overall score is conservative.

## Agentic Benchmarks

| Benchmark | Score |
|-----------|-------|
| [AutomationBench](/benchmarks/automationbench) | 36.1% |
| [Terminal-Bench-Science 0.1](/benchmarks/terminal-bench-science) | 57.0% |
| [ExploitGym](/benchmarks/exploitgym) | 35.1% |
| [GDPval-AA](/benchmarks/gdpvalaanormalized) | 53.8% |

## Coding Benchmarks

| Benchmark | Score |
|-----------|-------|
| [DeepSWE](/benchmarks/deepswe) | 71.9% |
| [AA-SciCode](/benchmarks/aascicode) | 54.2% |

## Multimodal & Grounded Benchmarks

| Benchmark | Score |
|-----------|-------|
| [AA-MMMU-Pro](/benchmarks/aammmupro) | 86.0% |

## Reasoning Benchmarks

| Benchmark | Score |
|-----------|-------|
| [AA-LCR](/benchmarks/lcr) | 83.0% |
| [CritPt](/benchmarks/critpt) | 31.7% |

## Knowledge Benchmarks

| Benchmark | Score |
|-----------|-------|
| [HealthBench (raw)](/benchmarks/healthbench) | 56.7% |
| [HealthBench (length-adjusted)](/benchmarks/healthbenchlengthadjusted) | 58.5% |
| [HealthBench Professional](/benchmarks/healthbenchprofessional) | 64.2% |
| [HealthBench Professional (raw)](/benchmarks/healthbenchprofessionalraw) | 67.2% |
| [HealthBench Hard](/benchmarks/healthbench-hard) | 36.2% |
| [Artificial Analysis Intelligence Index](/benchmarks/artificialanalysis) | 51.8% |
| [AA-HLE](/benchmarks/aahle) | 52.9% |
| [AA-Omniscience Index](/benchmarks/aaomniscienceindex) | 41.5% |
| [AA-Omniscience Accuracy](/benchmarks/omniscienceaccuracy) | 62.1% |
| [AA-Omniscience Hallucination Rate](/benchmarks/omnisciencehallucinationrate) | 54.3% |

## Other OpenAI Models

- [GPT-6 Astra](/models/gpt-6-astra) - Score: 88.22
- [GPT-6 Sol](/models/gpt-6-sol) - Score: 78.57
- [GPT-5.6 Sol](/models/gpt-5-6-sol) - Score: 78.19
- [GPT-5.6 Terra](/models/gpt-5-6-terra) - Score: 72.4
- [GPT-5.5 Pro](/models/gpt-5-5-pro) - Score: 72.18
- [GPT-5.4 Pro](/models/gpt-5-4-pro) - Score: 70.67
- [GPT-5.5](/models/gpt-5-5) - Score: 68.76
- [GPT-5.4](/models/gpt-5-4) - Score: 68.22
- [GPT-5.2 Pro](/models/gpt-5-2-pro) - Score: 66.74
- [GPT-6 Luna](/models/gpt-6-luna) - Score: 66.07
- [GPT-5.6 Luna](/models/gpt-5-6-luna) - Score: 65.39
- [GPT-5.3 Codex](/models/gpt-5-3-codex) - Score: 61.81
- [GPT-5.2](/models/gpt-5-2) - Score: 61.18
- [GPT-5.1](/models/gpt-5-1) - Score: 58.17
- [GPT-5.4 mini](/models/gpt-5-4-mini) - Score: 54.85
- [GPT-5.4 nano](/models/gpt-5-4-nano) - Score: 50.7
- [GPT-5.2-Codex](/models/gpt-5-2-codex) - Score: 49.86
- [o3-pro](/models/o3-pro) - Score: 47.13
- [o3](/models/o3) - Score: 46.95
- [GPT-5 (medium)](/models/gpt-5-medium) - Score: 44.64
- [GPT-5.1-Codex](/models/gpt-5-1-codex) - Score: 44.38
- [GPT-5 mini](/models/gpt-5-mini) - Score: 43.98
- [o3-mini](/models/o3-mini) - Score: 40.11
- [GPT-4.1](/models/gpt-4-1) - Score: 39.74
- [GPT-OSS 120B](/models/gpt-oss-120b) - Score: 37.66
- [o1](/models/o1) - Score: 37.48
- [o1-preview](/models/o1-preview) - Score: 36.25
- [o1-pro](/models/o1-pro) - Score: 35.7
- [GPT-5 nano](/models/gpt-5-nano) - Score: 34.96
- [GPT-OSS 20B](/models/gpt-oss-20b) - Score: 33.57
- [GPT-4o](/models/gpt-4o) - Score: 30.74
- [GPT-4.1 mini](/models/gpt-4-1-mini) - Score: 29.09
- [GPT-4o mini](/models/gpt-4o-mini) - Score: 26.87
- [GPT-4.1 nano](/models/gpt-4-1-nano) - Score: 24.89
- [GPT-4 Turbo](/models/gpt-4-turbo) - Score: 21.64
- [GPT-5.1-Codex-Max](/models/gpt-5-1-codex-max) - Score: not computed
- [o4-mini (high)](/models/o4-mini-high) - Score: not computed
- [GPT-5.2 Instant](/models/gpt-5-2-instant) - Score: not computed
- [GPT-5.3 Instant](/models/gpt-5-3-instant) - Score: not computed
- [GPT-5 (high)](/models/gpt-5-high) - Score: not computed
- [GPT-5.3-Codex-Spark](/models/gpt-5-3-codex-spark) - Score: not computed
- [GPT-Live-1](/models/gpt-live-1) - Score: not computed
- [GPT-5.6 Cyber](/models/gpt-5-6-cyber) - Score: not computed
- [GPT Live Transcribe](/models/gpt-live-transcribe) - Score: not computed
- [GPT Transcribe](/models/gpt-transcribe) - Score: not computed
- [GPT Realtime 2.1](/models/gpt-realtime-2-1) - Score: not computed
- [GPT Realtime 2.1 Mini](/models/gpt-realtime-2-1-mini) - Score: not computed
- [GPT Audio 1.5](/models/gpt-audio-1-5) - Score: not computed
- [GPT Realtime 1.5](/models/gpt-realtime-1-5) - Score: not computed
- [GPT-4o mini TTS](/models/gpt-4o-mini-tts) - Score: not computed
- [GPT Realtime](/models/gpt-realtime) - Score: not computed
- [GPT Realtime 2](/models/gpt-realtime-2) - Score: not computed
- [GPT Realtime mini](/models/gpt-realtime-mini) - Score: not computed
- [GPT-4o Audio](/models/gpt-4o-audio) - Score: not computed
- [GPT-4o mini Audio](/models/gpt-4o-mini-audio) - Score: not computed
