Skip to main content
ModelScale

Agentic benchmark

Finance Agent v2 leaderboard

Every model the catalog carries a published Finance Agent v2 value for, ranked by that value.

CategoryAgentic
MeasureMean score across repeated runs
TasksFinancial analyst task categories
DifficultyProfessional expert-task agent workflow
Published byFinance Agent v2

Vals AI benchmark for realistic financial analyst agent tasks across qualitative analysis, quantitative analysis, market work, comparables, precedents, earnings, disclosure, and modeling.

Finance Agent v2 ranking

5 models with a published Finance Agent v2 value, ordered by that value, highest first. Models the source has not scored on this benchmark are not listed — they are unmeasured, not last.

Models ranked by their published Finance Agent v2 value
RankModelProviderMean score across repeated runs
1Gemini 3.8 FlashGoogle61.4
2Ling 3.0 Flash FinInclusionAI59.8
3Gemini 3.5 FlashGoogle57.9
4Muse Spark 1.1Meta57.2
5Claude Opus 4.8Anthropic53.9

Evidence key: Observed

Rows are ordered by the value Finance Agent v2 published, highest first. The source does not state whether a higher value is the better result, so this page does not either: for a benchmark that measures a rate of failure, read the table from the bottom.

All leaderboards · Agentic capability leaderboard