Skip to main content
ModelScale

Coding benchmark

React Native Evals leaderboard

Every model the catalog carries a published React Native Evals value for, ranked by that value.

CategoryCoding
MeasureFramework-specific app development evaluation
TasksReact Native app implementation tasks
DifficultyProduction mobile app engineering
Published byReact Native Evals

An open benchmark for AI coding agents on real-world React Native implementation tasks, emphasizing working app behavior, recommended architecture choices, and strict constraint adherence.

React Native Evals ranking

16 models with a published React Native Evals value, ordered by that value, highest first. Models the source has not scored on this benchmark are not listed — they are unmeasured, not last.

Models ranked by their published React Native Evals value
RankModelProviderFramework-specific app development evaluation
1Composer 2Cursor96.1
2Composer 2 FastCursor94.9
3GPT-5.4OpenAI85.3
4GPT-5.5OpenAI84.7
5Claude Opus 4.6Anthropic84.1
6Claude Opus 4.7Anthropic82.8
7Claude Sonnet 4.6Anthropic80.6
8Gemini 3.1 ProGoogle78.9
9Kimi K2.5Moonshot AI77.2
10Gemma 4 31BGoogle75.2
11GLM-5Z.AI74.8
12Grok 4xAI72.6
13GPT-OSS 120BOpenAI71.6
14DeepSeek V3.2DeepSeek71.5
15MiniMax M2.7MiniMax71.4
16GPT-OSS 20BOpenAI71

Evidence key: Observed

Rows are ordered by the value React Native Evals published, highest first. The source does not state whether a higher value is the better result, so this page does not either: for a benchmark that measures a rate of failure, read the table from the bottom.

All leaderboards · Coding capability leaderboard