# Mistral Small 4 Benchmark Scores & Performance

> Mistral Small 4 by Mistral scores 42.97/100 overall, ranking #153 out of 483 AI models.

Canonical page: https://benchlm.ai/models/mistral-small-4

Last updated: September 10, 2026

## Model Details

| Property | Value |
|----------|-------|
| Creator | Mistral |
| Source Type | Open Weight |
| Reasoning Type | Non-Reasoning |
| Context Window | 256K |
| Overall Score | 42.97/100 |
| Overall Rank | #153 of 483 |

## Family & Coverage

- Family: Mistral Small 4
- Variant: base
- Benchmarks covered: 16 of 435
- Sibling models: [Mistral Small 4 (Reasoning)](/models/mistral-small-4-reasoning)
- Coverage note: BenchLM currently has partial benchmark coverage for this model, so the overall score is conservative.

## Agentic Benchmarks

| Benchmark | Score |
|-----------|-------|
| [τ²-bench results](/benchmarks/tau2-bench) | 41.2% |
| [GDPval-AA](/benchmarks/gdpvalaanormalized) | 1.8% |
| [AA Agentic Index](/benchmarks/aaagenticindex) | 1.4% |
| [GDPval-AA](/benchmarks/gdpvalaa) | 537 |

## Coding Benchmarks

| Benchmark | Score |
|-----------|-------|
| [AA-SciCode](/benchmarks/aascicode) | 38.8% |
| [AA Coding Index](/benchmarks/aacodingindex) | 26.6% |

## Multimodal & Grounded Benchmarks

| Benchmark | Score |
|-----------|-------|
| [AA-MMMU-Pro](/benchmarks/aammmupro) | 56.8% |

## Reasoning Benchmarks

| Benchmark | Score |
|-----------|-------|
| [AA-LCR](/benchmarks/lcr) | 49.7% |
| [CritPt](/benchmarks/critpt) | 0.3% |

## Knowledge Benchmarks

| Benchmark | Score |
|-----------|-------|
| [Artificial Analysis Intelligence Index](/benchmarks/artificialanalysis) | 11.4% |
| [AA-GPQA Diamond](/benchmarks/aagpqadiamond) | 76.9% |
| [AA-HLE](/benchmarks/aahle) | 9.9% |
| [AA-Omniscience Index](/benchmarks/aaomniscienceindex) | -30.4% |
| [AA-Omniscience Accuracy](/benchmarks/omniscienceaccuracy) | 21.7% |
| [AA-Omniscience Hallucination Rate](/benchmarks/omnisciencehallucinationrate) | 66.5% |

## Instruction Following Benchmarks

| Benchmark | Score |
|-----------|-------|
| [AA-IFBench](/benchmarks/aaifbench) | 48.2% |

## Other Mistral Models

- [Ministral 3 14B (Reasoning)](/models/ministral-3-14b-reasoning) - Score: 47.18
- [Mistral Large 3](/models/mistral-large-3) - Score: 45.61
- [Mistral 8x7B](/models/mistral-8x7b) - Score: 42.91
- [Ministral 3 8B (Reasoning)](/models/ministral-3-8b-reasoning) - Score: 38.41
- [Ministral 3 3B (Reasoning)](/models/ministral-3-3b-reasoning) - Score: 37.56
- [Mistral 8x7B v0.2](/models/mistral-8x7b-v0-2) - Score: 37.19
- [Mistral Large 2](/models/mistral-large-2) - Score: 36.41
- [Mistral Medium 3](/models/mistral-medium-3) - Score: 30.55
- [Ministral 3 14B](/models/ministral-3-14b) - Score: 30.27
- [Mistral Medium 3.5 128B](/models/mistral-medium-3-5-128b) - Score: 30.13
- [Mixtral 8x22B Instruct v0.1](/models/mixtral-8x22b-instruct-v0-1) - Score: 26.65
- [Ministral 3 8B](/models/ministral-3-8b) - Score: 17.87
- [Ministral 3 3B](/models/ministral-3-3b) - Score: 16.04
- [Mistral 7B v0.3](/models/mistral-7b-v0-3) - Score: 8.66
- [Leanstral 1.5](/models/leanstral-1-5) - Score: not computed
- [Leanstral](/models/leanstral) - Score: not computed
- [Mistral Small 4 (Reasoning)](/models/mistral-small-4-reasoning) - Score: not computed
