Skip to main content
BenchLM

Kimi Code Bench v2

We show this table for reference; we do not rank on it.

Data verified 34 confirmed releases in the last 30 daysFollow model changes

A Moonshot AI internal coding-agent benchmark for realistic software-engineering tasks across mainstream programming languages and production technology stacks.

Benchmark score on Kimi Code Bench v2 — September 27, 2026

We compile the Kimi Code Bench v2 rows from provider self-reports. Kimi K3 leads the table at 72.9%, followed by Kimi K2.7 Code (62.0%). We do not use these results to rank models overall.

2 modelsCodingCurrentDisplay onlyUpdated September 27, 2026

Benchmark score table (2 models)

Score
1
Kimi K3Moonshot AI · Closed
72.9%
2
Kimi K2.7 CodeMoonshot AI · Open weight
62.0%

About Kimi Code Bench v2

Year

2026

Tasks

Realistic coding-agent tasks

Format

Coding-agent pass rate

Difficulty

Production software engineering

Moonshot describes Kimi Code Bench v2 as an in-house coding-agent benchmark covering backend services, infrastructure, performance engineering, systems programming, security, frontend development, and ML/data engineering. BenchLM stores provider-reported exact values as display-only launch evidence.

Freshness and provenance

Version

Kimi Code Bench v2 2026

Refresh cadence

Quarterly

Staleness state

Current

Question availability

Public benchmark set

CurrentDisplay only

BenchLM uses freshness metadata to decide whether a benchmark should still be treated as a strong differentiator, a benchmark to watch, or a display-only reference. For the full scoring policy, see the BenchLM methodology page.

Questions

What does Kimi Code Bench v2 measure?

A Moonshot AI internal coding-agent benchmark for realistic software-engineering tasks across mainstream programming languages and production technology stacks.

Which model scores highest on Kimi Code Bench v2?

Kimi K3 by Moonshot AI currently leads with a score of 72.9% on Kimi Code Bench v2.

How many models are evaluated on Kimi Code Bench v2?

2 AI models have been evaluated on Kimi Code Bench v2 on BenchLM.

Compare top models on Kimi Code Bench v2

Last updated: September 27, 2026 · BenchLM version Kimi Code Bench v2 2026

Know when it’s worth switching models

The model to choose, the cheaper alternative, and the release we would wait on.

Read a sample issue

Join 2,000+ readers.

One email each week. Unsubscribe anytime.