# MMMU-Pro with Python (MMMU-Pro w/ Python)

> Tool-augmented MMMU-Pro variant that allows Python assistance during multimodal reasoning.

Canonical page: https://benchlm.ai/benchmarks/mmmupropython

- Category: [Multimodal & Grounded](/multimodal-grounded)
- Last updated: September 27, 2026

## About MMMU-Pro w/ Python

- Year: 2026
- Tasks: Multimodal academic reasoning
- Format: Image + text question answering with Python
- Difficulty: Frontier multimodal
- Paper: [Introducing GPT-5.4 mini and nano](https://openai.com/index/introducing-gpt-5-4-mini-and-nano/)

Useful for measuring multimodal reasoning when the model can combine visual understanding with computation.

MMMU-Pro w/ Python is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Leaderboard (9 models)

| Rank | Model | Creator | Score |
|------|-------|---------|-------|
| 1 | [GPT-5.6 Sol](/models/gpt-5-6-sol) | OpenAI | 84.6% |
| 2 | [Kimi K3](/models/kimi-k3) | Moonshot AI | 83.4% |
| 3 | [GPT-5.5](/models/gpt-5-5) | OpenAI | 83.2% |
| 4 | [GPT-5.4](/models/gpt-5-4) | OpenAI | 82.1% |
| 5 | [GPT-5.6 Terra](/models/gpt-5-6-terra) | OpenAI | 82% |
| 6 | [Kimi K2.6](/models/kimi-2-6) | Moonshot AI | 80.1% |
| 7 | [GPT-5.6 Luna](/models/gpt-5-6-luna) | OpenAI | 79.5% |
| 8 | [GPT-5.4 mini](/models/gpt-5-4-mini) | OpenAI | 78% |
| 9 | [GPT-5.4 nano](/models/gpt-5-4-nano) | OpenAI | 69.5% |

## FAQ

### What does MMMU-Pro w/ Python measure?

Tool-augmented MMMU-Pro variant that allows Python assistance during multimodal reasoning.

### Which model scores highest on MMMU-Pro w/ Python?

GPT-5.6 Sol by OpenAI currently leads with a score of 84.6% on MMMU-Pro w/ Python.

### How many models are evaluated on MMMU-Pro w/ Python?

9 AI models have been evaluated on MMMU-Pro w/ Python on BenchLM.

### Does MMMU-Pro w/ Python affect BenchLM's overall score?

Not directly. MMMU-Pro w/ Python is still displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Compare Top Models on MMMU-Pro w/ Python

- [GPT-5.6 Sol vs Kimi K3](/compare/gpt-5-6-sol-vs-kimi-k3)
- [Kimi K3 vs GPT-5.5](/compare/gpt-5-5-vs-kimi-k3)
- [GPT-5.5 vs GPT-5.4](/compare/gpt-5-4-vs-gpt-5-5)
- [GPT-5.4 vs GPT-5.6 Terra](/compare/gpt-5-4-vs-gpt-5-6-terra)
