Skip to main content
ModelScale

Coding benchmark

AA Coding Agents leaderboard

Artificial Analysis Coding Agent Index. Every model the catalog carries a published AA Coding Agents value for, ranked by that value.

CategoryCoding
MeasureAverage pass@1 index
TasksComposite over DeepSWE, Terminal-Bench v2, and SWE-Atlas-QnA
DifficultyReal-world coding-agent workflows

A display-only Artificial Analysis leaderboard for coding-agent systems, combining agent harnesses, host models, and execution settings across software-engineering benchmarks.

No model in the catalog has a published AA Coding Agents score.

The benchmark is defined by Artificial Analysis Coding Agent Benchmarks, but the catalog carries no value for it yet. An absent value is shown as absent here rather than as a zero.

All leaderboards · Coding capability leaderboard