# SpreadsheetBench 2

> A spreadsheet-focused benchmark for agentic analysis and editing workflows.

Canonical page: https://benchlm.ai/benchmarks/spreadsheetbench2

- Category: [Agentic](/agentic)
- Last updated: September 10, 2026

## About SpreadsheetBench 2

- Year: 2026
- Tasks: Spreadsheet analysis and editing tasks
- Format: Agent task-completion score
- Difficulty: Professional spreadsheet work
- Paper: [Kimi K3: Open Frontier Intelligence](https://www.kimi.com/blog/kimi-k3)

Moonshot evaluates Kimi K3 with the Claude Code harness. BenchLM stores the provider-exact value as display-only productivity evidence.

SpreadsheetBench 2 is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Leaderboard (2 models)

| Rank | Model | Creator | Score |
|------|-------|---------|-------|
| 1 | [Kimi K3](/models/kimi-k3) | Moonshot AI | 34.8% |
| 2 | [Ling 3.0 Flash Fin](/models/ling-3-0-flash-fin) | InclusionAI | 21.8% |

## FAQ

### What does SpreadsheetBench 2 measure?

A spreadsheet-focused benchmark for agentic analysis and editing workflows.

### Which model scores highest on SpreadsheetBench 2?

Kimi K3 by Moonshot AI currently leads with a score of 34.8% on SpreadsheetBench 2.

### How many models are evaluated on SpreadsheetBench 2?

2 AI models have been evaluated on SpreadsheetBench 2 on BenchLM.

### Does SpreadsheetBench 2 affect BenchLM's overall score?

Not directly. SpreadsheetBench 2 is still displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Compare Top Models on SpreadsheetBench 2

- [Kimi K3 vs Ling 3.0 Flash Fin](/compare/kimi-k3-vs-ling-3-0-flash-fin)
