Make it yours

Make a custom ranking reflect your deployment priorities. Capability evidence drives the score; provider access, runtime constraints, and token cost remain separate.

Published custom ranking unavailable

No validated publication is available.

Capability weighting matrix

Five capability weights form the quality score. Runtime remains a separate threshold.

90 quality weight points · coverage and formula

Composite = Σ(derived rank-percentile × entered weight) / Σ(active weights). The entered total is the denominator; weights are not silently redistributed. Runtime is excluded from the quality score.

  • Agentic: no comparable evidence published
  • Coding: no comparable evidence published
  • Reasoning: no comparable evidence published
  • Math: no comparable evidence published
  • Multimodal: no comparable evidence published
  • Throughput: runtime threshold only, excluded from quality weights
Positive-weight axes without published comparable evidence: Agentic, Coding, Reasoning, Math, Multimodal.

Set these axes to zero, or explicitly apply equal weights to published axes.

Model access
Providers · all
Filter providers

Provider evidence unavailable.

0 / 4 selected

Weighted ranking

0 ranked · 0 lack one or more positively weighted axes · 0 have unobserved runtime.

Ranking and export scope

Chart: top 20 plus explicitly selected eligible models. PNG exports both charts with weights and source context; CSV includes every matching row.

No eligible scores for these weights.

Observed runtime constraints

Absent runtime measurements are unobserved. Hiding unobserved/outside-SLA rows requires both qualified measurements to meet the selected thresholds.

Observed runtime SLA

TTFT (seconds)

No qualified runtime observations are published. Unknown runtime is not a measured failure.

Output speed (tok/s)

No qualified runtime observations are published. Unknown runtime is not a measured failure.

Exact SLA measurements
ModelTTFT secondsTTFT assessmentThroughput tok/sThroughput assessmentCombined SLA

Weighted score vs. cost

Prices shown were recorded with these benchmark results; current prices may differ. Exact route and catalog revision are included in the source receipt.

Blended price = (3 × input price + output price) / 4, in USD per 1M tokens: an explicit 3:1 input/output assumption, excluding cache, batch and long-context adjustments. The blend is rounded upward to one microUSD per 1M tokens (less than $0.000001 above the exact blend); source integer rate precision is preserved in the receipt. 0 ranked rows lack a verified route with both token prices. These models remain eligible for quality ranking; their cost frontier is unavailable.

Score frontier

Verified input and output token prices are not reported for these rows. Quality ranking remains available independently.

Cheapest-first score ranking

No eligible scores for these weights.

Exact score and token-cost values (cheapest 20)
Cost orderModelProviderWeighted scoreBlended $ / 1M tokens USDFrontier

Ranked output

Exact values use the same ordered rows as the score chart and exported receipt.

Select / rankModelProviderWeightedBlended $ / 1MInput $ / 1MOutput $ / 1MTTFT sThroughput tok/sSLAFrontier
No models satisfy this combination of published capability evidence and filters. Missing evidence is never scored as zero.

Methodology & source receipt

Rank-percentiles are derived from a published rank and exact cohort: 100 × (cohort size − rank) / (cohort size − 1). Only matching published capability definitions are compared. Agentic, Coding, Reasoning, Math and Multimodal each require their own exact source binding. Throughput is a separately qualified runtime measurement and never enters the quality score.

Exact weights, axis definitions and publication lineage
{
  "weights": {
    "agentic": 20,
    "coding": 20,
    "reasoning": 20,
    "mathematics": 15,
    "multimodal": 15
  },
  "revision": null,
  "cacheRevision": null,
  "tokenCostBlend": "3:1 input/output; ceiling to microUSD per 1M tokens",
  "benchmarkCatalogRevision": null,
  "currentCatalogRevision": null,
  "axes": null
}
Current output page: per-model capability, price and runtime conditions
[]