# Gemini 3.7 Flash Benchmark Scores & Performance

> Gemini 3.7 Flash by Google scores 67.66/100 overall, ranking #19 out of 508 AI models.

Canonical page: https://benchlm.ai/models/gemini-3-7-flash

Last updated: September 27, 2026

## Model Details

| Property | Value |
|----------|-------|
| Creator | Google |
| Source Type | Proprietary |
| Reasoning Type | Reasoning |
| Context Window | 1M |
| Official model card | [Google DeepMind Gemini 3.7 Flash model card](https://deepmind.google/models/model-cards/gemini-3-7-flash/) |
| Overall Score | 67.66/100 |
| Overall Rank | #19 of 508 |

## Family & Coverage

- Family: Gemini 3.7 Flash
- Variant: base
- Benchmarks covered: 42 of 486
- Related earlier model: [Gemini 3.6 Flash](/models/gemini-3-6-flash)
- Coverage note: BenchLM currently has partial benchmark coverage for this model, so the overall score is conservative.

## Agentic Benchmarks

| Benchmark | Score |
|-----------|-------|
| [Terminal-Bench 2.1](/benchmarks/terminalbench21) | 85.8% |
| [Terminal-Bench 3.0](/benchmarks/terminal-bench-3) | 14.9% |
| [AutomationBench](/benchmarks/automationbench) | 30.4% |
| [OSWorld 2.0](/benchmarks/osworld2) | 47.9% |
| [Agents' Last Exam](/benchmarks/agentslastexam) | 26.3% |
| [GDPval-AA](/benchmarks/gdpvalaanormalized) | 43.6% |
| [GDPval-AA](/benchmarks/gdpvalaa) | 1525 |
| [AA Harvey LAB](/benchmarks/aaharveylab) | 90.7% |
| [Terminal-Bench 2.1 (Vals)](/benchmarks/valsterminalbench21) | 77.5% |
| [AA Agentic Index](/benchmarks/aaagenticindex) | 36.4% |
| [AA-AnalystAgent](/benchmarks/aaanalystagent) | 60.0% |
| [ApprenticeBench](/benchmarks/apprenticebench) | 16% |

## Coding Benchmarks

| Benchmark | Score |
|-----------|-------|
| [FrontierCode 1.1 Main](/benchmarks/frontiercode) | 43.6% |
| [DeepSWE](/benchmarks/deepswe) | 65.3% |
| [Terminal-Bench 2.1](/benchmarks/terminalbench21) | 85.8% |
| [AA-SciCode](/benchmarks/aascicode) | 57.2% |
| [FrontierSWE v2](/benchmarks/frontierswev2) | 20.3% |
| [LiveCodeBench (Vals)](/benchmarks/valslivecodebench) | 88.7% |
| [SWE-bench (Vals)](/benchmarks/valsswebench) | 80.8% |
| [AA Coding Index](/benchmarks/aacodingindex) | 76.1% |

## Multimodal & Grounded Benchmarks

| Benchmark | Score |
|-----------|-------|
| [CharXiv w/o tools](/benchmarks/charxivnotools) | 84.5% |
| [CharXiv](/benchmarks/charxiv) | 88.7% |
| [LVBench](/benchmarks/lvbench) | 85.4% |
| [AA-MMMU-Pro](/benchmarks/aammmupro) | 85.5% |
| [Design Arena Website](/benchmarks/designarenawebsite) | 1313 |

## Reasoning Benchmarks

| Benchmark | Score |
|-----------|-------|
| [MRCR v2 64K-128K](/benchmarks/mrcr-v2-64k-128k) | 97% |
| [AA-LCR](/benchmarks/lcr) | 81.7% |
| [CritPt](/benchmarks/critpt) | 14.3% |
| [ARC-AGI-1](/benchmarks/arcagi1) | 95.50% |
| [ARC-AGI-2](/benchmarks/arc-agi-2) | 84.6% |

## Knowledge Benchmarks

| Benchmark | Score |
|-----------|-------|
| [HLE-Verified](/benchmarks/hleverified) | 53.6% |
| [LABBench2](/benchmarks/labbench2) | 82.1% |
| [BioMysteryBench (human-solvable)](/benchmarks/biomysterybenchhumansolvable) | 87.1% |
| [BioMysteryBench (human-difficult)](/benchmarks/biomysterybenchhumandifficult) | 43.5% |
| [Artificial Analysis Intelligence Index](/benchmarks/artificialanalysis) | 39.1% |
| [AA-GPQA Diamond](/benchmarks/aagpqadiamond) | 94.5% |
| [AA-HLE](/benchmarks/aahle) | 47.9% |
| [AA-Omniscience Index](/benchmarks/aaomniscienceindex) | 26.5% |
| [AA-Omniscience Accuracy](/benchmarks/omniscienceaccuracy) | 55.3% |
| [AA-Omniscience Hallucination Rate](/benchmarks/omnisciencehallucinationrate) | 64.5% |
| [GPQA Diamond (Vals)](/benchmarks/valsgpqadiamond) | 93.9% |
| [MMLU-Pro (Vals)](/benchmarks/valsmmlupro) | 90.1% |

## Other Google Models

- [Gemini 3.8 Flash](/models/gemini-3-8-flash) - Score: 73.43
- [Gemini 3.6 Flash](/models/gemini-3-6-flash) - Score: 64.58
- [Gemini 3.1 Pro](/models/gemini-3-1-pro) - Score: 64.45
- [Gemini 3.5 Flash](/models/gemini-3-5-flash) - Score: 63.49
- [Gemini 3 Pro](/models/gemini-3-pro) - Score: 61.01
- [Gemini 3 Flash](/models/gemini-3-flash) - Score: 55.61
- [Gemini 3.5 Flash-Lite](/models/gemini-3-5-flash-lite) - Score: 50.96
- [Gemini 2.5 Pro](/models/gemini-2-5-pro) - Score: 50.24
- [Gemini 3.1 Flash-Lite](/models/gemini-3-1-flash-lite) - Score: 48.51
- [Gemma 4 26B A4B](/models/gemma-4-26b-a4b) - Score: 46.28
- [Gemma 4 31B](/models/gemma-4-31b) - Score: 44.84
- [Gemini 2.5 Flash](/models/gemini-2-5-flash) - Score: 43
- [Gemma 4 E4B](/models/gemma-4-e4b) - Score: 33.1
- [Gemma 4 E2B](/models/gemma-4-e2b) - Score: 32.27
- [Gemma 3 27B](/models/gemma-3-27b) - Score: 28.74
- [Gemma 4 12B](/models/gemma-4-12b) - Score: 28.47
- [Gemini 1.5 Pro](/models/gemini-1-5-pro) - Score: 27.93
- [Gemini 1.0 Pro](/models/gemini-1-0-pro) - Score: 12.95
- [Gemini 3 Pro Deep Think](/models/gemini-3-pro-deep-think) - Score: not computed
- [Gemini 3.8 Flash TTS](/models/gemini-3-8-flash-tts) - Score: not computed
- [Gemini 3.8 Flash-Lite TTS](/models/gemini-3-8-flash-lite-tts) - Score: not computed
- [Gemini 3.8 Live](/models/gemini-3-8-live) - Score: not computed
- [Gemini 3.8 Live Extended Thinking](/models/gemini-3-8-live-extended-thinking) - Score: not computed
- [Gemini 3.8 Flash Cyber](/models/gemini-3-8-flash-cyber) - Score: not computed
- [Gemini 3.5 Transcribe](/models/gemini-3-5-transcribe) - Score: not computed
- [Gemini 3.5 Flash Cyber](/models/gemini-3-5-flash-cyber) - Score: not computed
- [Gemini 3.1 Flash TTS Preview](/models/gemini-3-1-flash-tts-preview) - Score: not computed
- [Gemini 2.5 Flash Native Audio Preview (12-2025)](/models/gemini-2-5-flash-native-audio-preview-12-2025) - Score: not computed
- [Gemini 2.5 Flash TTS Preview](/models/gemini-2-5-flash-tts-preview) - Score: not computed
- [Gemini 2.5 Pro TTS Preview](/models/gemini-2-5-pro-tts-preview) - Score: not computed
- [Gemini 3.1 Flash Live Preview](/models/gemini-3-1-flash-live-preview) - Score: not computed
