# Kimi Code Bench v2

> A Moonshot AI internal coding-agent benchmark for realistic software-engineering tasks across mainstream programming languages and production technology stacks.

Canonical page: https://benchlm.ai/benchmarks/kimicodebenchv2

- Category: [Coding](/coding)
- Last updated: September 10, 2026

## About Kimi Code Bench v2

- Year: 2026
- Tasks: Realistic coding-agent tasks
- Format: Coding-agent pass rate
- Difficulty: Production software engineering
- Paper: [Kimi K2.7 Code](https://huggingface.co/moonshotai/Kimi-K2.7-Code)

Moonshot describes Kimi Code Bench v2 as an in-house coding-agent benchmark covering backend services, infrastructure, performance engineering, systems programming, security, frontend development, and ML/data engineering. BenchLM stores provider-reported exact values as display-only launch evidence.

Kimi Code Bench v2 is currently displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Leaderboard (2 models)

| Rank | Model | Creator | Score |
|------|-------|---------|-------|
| 1 | [Kimi K3](/models/kimi-k3) | Moonshot AI | 72.9% |
| 2 | [Kimi K2.7 Code](/models/kimi-k2-7-code) | Moonshot AI | 62.0% |

## FAQ

### What does Kimi Code Bench v2 measure?

A Moonshot AI internal coding-agent benchmark for realistic software-engineering tasks across mainstream programming languages and production technology stacks.

### Which model scores highest on Kimi Code Bench v2?

Kimi K3 by Moonshot AI currently leads with a score of 72.9% on Kimi Code Bench v2.

### How many models are evaluated on Kimi Code Bench v2?

2 AI models have been evaluated on Kimi Code Bench v2 on BenchLM.

### Does Kimi Code Bench v2 affect BenchLM's overall score?

Not directly. Kimi Code Bench v2 is still displayed on BenchLM for reference, but it is excluded from the weighted scoring formula.

## Compare Top Models on Kimi Code Bench v2

- [Kimi K3 vs Kimi K2.7 Code](/compare/kimi-k2-7-code-vs-kimi-k3)
