ModelRefs / MLDR Leaderboard — AI Model Scores
MLDR Leaderboard — AI Model Scores
Multilingual long-document retrieval benchmark introduced with M3-Embedding for retrieval across 13 languages. Current leaders, methodology, and citation sources for MLDR.
Overview
Multilingual long-document retrieval benchmark introduced with M3-Embedding for retrieval across 13 languages.
How it is measured: Aggregate nDCG@10 over the multilingual long-document retrieval test sets; model records must preserve the reported dense, sparse, or hybrid configuration.
How this benchmark is scored
| Category | retrieval |
|---|---|
| Maximum score | 100 average nDCG@10 |
| Direction | Higher is better |
Primary source: https://arxiv.org/abs/2402.03216
Continue your research
Use these connected ModelRefs sections to compare alternatives, inspect implementation paths, and review the evidence and governance boundaries relevant to MLDR Leaderboard — AI Model Scores.