# Vals Multimodal Index v1.2 (Vals Multimodal Index)

> Vals AI multimodal composite across finance, coding, education, and mortgage-tax task families.

Canonical page: https://benchlm.ai/benchmarks/valsmultimodalindex

- Category: [Multimodal & Grounded](/multimodal-grounded)
- Last updated: August 11, 2026

## About Vals Multimodal Index

- Year: 2026
- Tasks: Finance, coding, education, and mortgage-tax components
- Format: Composite score
- Difficulty: Private multimodal economic-work benchmark composite
- Paper: [Vals Multimodal Index](https://www.vals.ai/benchmarks/vals_multimodal_index)

BenchLM mirrors the Vals Multimodal Index as a display-only external composite with task-level component scores preserved in the Vals snapshot.

Vals Multimodal Index is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Leaderboard (33 models)

| Rank | Model | Configuration | Creator | Score |
|------|-------|---------------|---------|-------|
| 1 | [Claude Fable 5](/models/claude-fable) | — | Anthropic | 74.15% |
| 2 | [Claude Opus 5](/models/claude-opus-5) | — | Anthropic | 73.90% |
| 3 | [Kimi K3](/models/kimi-k3) | — | Moonshot AI | 73.42% |
| 4 | [GPT-5.6 Sol](/models/gpt-5-6-sol) | max reasoning | OpenAI | 72.64% |
| 5 | [Claude Opus 4.8](/models/claude-opus-4-8) | — | Anthropic | 70.89% |
| 6 | [Muse Spark 1.2](/models/muse-spark-1-2) | — | Meta | 69.80% |
| 7 | [GPT-5.6 Luna](/models/gpt-5-6-luna) | max reasoning | OpenAI | 69.06% |
| 8 | [Claude Sonnet 5](/models/claude-sonnet-5) | — | Anthropic | 68.83% |
| 9 | [GPT-5.5](/models/gpt-5-5) | xhigh reasoning | OpenAI | 68.07% |
| 10 | [Claude Opus 4.7](/models/claude-opus-4-7) | — | Anthropic | 67.36% |
| 11 | [Muse Spark 1.1](/models/muse-spark-1-1) | xhigh reasoning | Meta | 66.74% |
| 12 | [Qwen3.8 Max](/models/qwen3-8-max) | — | Alibaba | 65.39% |
| 13 | [Gemini 3.6 Flash](/models/gemini-3-6-flash) | high reasoning | Google | 65.08% |
| 14 | [GPT-5.6 Terra](/models/gpt-5-6-terra) | xhigh reasoning | OpenAI | 65.07% |
| 15 | [Grok 4.5](/models/grok-4-5) | high reasoning | xAI | 63.42% |
| 16 | [Gemini 3.5 Flash](/models/gemini-3-5-flash) | high reasoning | Google | 62.93% |
| 17 | [Claude Sonnet 4.6](/models/claude-sonnet-4-6) | — | Anthropic | 60.57% |
| 18 | [MiniMax M3](/models/minimax-m3) | — | MiniMax | 59.97% |
| 19 | [Kimi K2.6](/models/kimi-2-6) | — | Moonshot AI | 56.43% |
| 20 | [Gemini 3.1 Pro Preview](https://www.vals.ai/models/google_gemini-3.1-pro-preview) | high reasoning | Google | 56.07% |
| 21 | [GPT-5.4 mini](/models/gpt-5-4-mini) | xhigh reasoning | OpenAI | 54.22% |
| 22 | [Qwen3.7 Plus](/models/qwen3-7-plus) | — | Alibaba | 53.89% |
| 23 | [MiMo-V2.5](/models/mimo-v2-5) | — | Xiaomi | 52.77% |
| 24 | [Gemini 3 Flash Preview](https://www.vals.ai/models/google_gemini-3-flash-preview) | high reasoning | Google | 52.19% |
| 25 | [Qwen3.6 Plus](/models/qwen3-6-plus) | — | Alibaba | 51.52% |
| 26 | [Inkling-Small](/models/inkling-small) | 0.99 reasoning | Thinking Machines Lab | 50.05% |
| 27 | [Inkling](/models/inkling) | 0.99 reasoning | Thinking Machines Lab | 49.40% |
| 28 | [GPT-5.4 nano](/models/gpt-5-4-nano) | high reasoning | OpenAI | 47.64% |
| 29 | [Grok 4.3](/models/grok-4-3) | high reasoning | xAI | 43.29% |
| 30 | [Claude Haiku 4.5 Thinking](/models/claude-haiku-4-5-thinking) | — | Anthropic | 42.88% |
| 31 | [Gemini 3.1 Flash Lite Preview](https://www.vals.ai/models/google_gemini-3.1-flash-lite-preview) | high reasoning | Google | 41.35% |
| 32 | [Grok 4.20 0309 Reasoning](https://www.vals.ai/models/grok_grok-4.20-0309-reasoning) | — | xAI | 39.06% |
| 33 | [Mistral Medium 3.5](https://www.vals.ai/models/mistralai_mistral-medium-3.5) | high reasoning | Mistral AI | 34.77% |

## FAQ

### What does Vals Multimodal Index measure?

Vals AI multimodal composite across finance, coding, education, and mortgage-tax task families.

### Which model leads the published Vals Multimodal Index snapshot?

Claude Fable 5 currently leads the published Vals Multimodal Index snapshot with a score of 74.15%.

### How many models are evaluated on Vals Multimodal Index?

The August 11, 2026 contains 33 AI models.

### Does Vals Multimodal Index affect BenchLM's overall score?

Not directly. Vals Multimodal Index is still displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.
