Skip to main content
BenchLM
Data

React Native Evals

We show this table for reference; we do not rank on it.

Data verified 37 confirmed releases in the last 30 daysFollow model changes

An open benchmark for AI coding agents on real-world React Native implementation tasks, emphasizing working app behavior, recommended architecture choices, and strict constraint adherence.

Benchmark score on React Native Evals — October 6, 2026

We compile the React Native Evals rows from benchmark-owner or independent runs. Composer 2 leads the table at 96.1%, followed by Composer 2 Fast (94.9%) and GPT-5.4 (85.3%). We do not use these results to rank models overall.

16 modelsCodingCurrentDisplay onlyUpdated October 6, 2026

Benchmark score results for React Native Evals
RankModel / configurationScoreParameters (B)Open / closed
1Composer 2Cursor
96.1%
Not reportedClosed
2Composer 2 FastCursor
94.9%
Not reportedClosed
3GPT-5.4OpenAI
85.3%
Not reportedClosed
4GPT-5.5OpenAI
84.7%
Not reportedClosed
5Claude Opus 4.6Anthropic
84.1%
Not reportedClosed
6Claude Opus 4.7Anthropic
82.8%
Not reportedClosed
7Claude Sonnet 4.6Anthropic
80.6%
Not reportedClosed
8Gemini 3.1 ProGoogle
78.9%
Not reportedClosed
9Kimi K2.5Moonshot AI
77.2%
Not reportedOpen
10Gemma 4 31BGoogle
75.2%
Not reportedOpen
11GLM-5Z.AI
74.8%
Not reportedOpen
12Grok 4xAI
72.6%
Not reportedClosed
13GPT-OSS 120BOpenAI
71.6%
Not reportedOpen
14DeepSeek V3.2DeepSeek
71.5%
Not reportedOpen
15MiniMax M2.7MiniMax
71.4%
Not reportedOpen
16GPT-OSS 20BOpenAI
71%
Not reportedOpen

Among the reported React Native Evals rows, Composer 2 is first at 96.1%. The third row is 10.8 points behind. The broader top-10 range is 20.9 points, so the table still separates the published systems.

16 models have been evaluated on React Native Evals. The benchmark falls in the Coding category. React Native Evals is currently displayed for reference but excluded from the scoring formula, so it does not directly affect overall rankings.

About React Native Evals

Year

2026

Tasks

React Native app implementation tasks

Format

Framework-specific app development evaluation

Difficulty

Production mobile app engineering

React Native Evals focuses on framework-specific mobile work that generic coding benchmarks often miss. The public dashboard groups tasks into areas like navigation, animation, and async state, with repeated runs and cost tracking across models.

Freshness and provenance

Version

React Native Evals 2026

Refresh cadence

Quarterly

Staleness state

Current

Question availability

Public benchmark set

CurrentDisplay only

BenchLM uses freshness metadata to decide whether a benchmark should still be treated as a strong differentiator, a benchmark to watch, or a display-only reference. For the full scoring policy, see the BenchLM methodology page.

Questions

What does React Native Evals measure?

An open benchmark for AI coding agents on real-world React Native implementation tasks, emphasizing working app behavior, recommended architecture choices, and strict constraint adherence.

Which model scores highest on React Native Evals?

Composer 2 by Cursor currently leads with a score of 96.1% on React Native Evals.

How many models are evaluated on React Native Evals?

16 AI models have published results on React Native Evals in the BenchLM catalog.

Last updated: October 6, 2026 · BenchLM version React Native Evals 2026

Know when it’s worth switching models

The model to choose, the cheaper alternative, and the release we would wait on.

Read a sample issue

Join 5,500+ readers.

One email each week. Unsubscribe anytime.