Skip to main content
ModelScale

Multilingual benchmark

MKQA-11 leaderboard

MKQA-11 multilingual retrieval. Every model the catalog carries a published MKQA-11 value for, ranked by that value.

CategoryMultilingual
MeasureRecall@20 average
TasksCross-lingual open-domain QA retrieval
DifficultyMultilingual retrieval

A display-only multilingual QA retrieval benchmark reported by Liquid AI for LFM2.5 retriever models, using Recall@20 across 11 languages.

MKQA-11 ranking

2 models with a published MKQA-11 value, ordered by that value, highest first. Models the source has not scored on this benchmark are not listed — they are unmeasured, not last.

Models ranked by their published MKQA-11 multilingual retrieval value
RankModelProviderRecall@20 average
1LFM2.5-ColBERT-350MLiquidAI69.4
2LFM2.5-Embedding-350MLiquidAI69.1

Evidence key: Observed

Rows are ordered by the value LFM2.5 Retrievers: Bi-directional LFMs for Fast Multilingual Search published, highest first. The source does not state whether a higher value is the better result, so this page does not either: for a benchmark that measures a rate of failure, read the table from the bottom.

All leaderboards