Coding benchmark
AA Coding Agents leaderboard
Artificial Analysis Coding Agent Index. Every model the catalog carries a published AA Coding Agents value for, ranked by that value.
CategoryCoding
MeasureAverage pass@1 index
TasksComposite over DeepSWE, Terminal-Bench v2, and SWE-Atlas-QnA
DifficultyReal-world coding-agent workflows
Published byArtificial Analysis Coding Agent Benchmarks
A display-only Artificial Analysis leaderboard for coding-agent systems, combining agent harnesses, host models, and execution settings across software-engineering benchmarks.
No model in the catalog has a published AA Coding Agents score.
The benchmark is defined by Artificial Analysis Coding Agent Benchmarks, but the catalog carries no value for it yet. An absent value is shown as absent here rather than as a zero.