ModelRefs / MLDR Leaderboard — AI Model Scores

MLDR Leaderboard — AI Model Scores

Multilingual long-document retrieval benchmark introduced with M3-Embedding for retrieval across 13 languages. Current leaders, methodology, and citation sources for MLDR.

Overview

Multilingual long-document retrieval benchmark introduced with M3-Embedding for retrieval across 13 languages.

How it is measured: Aggregate nDCG@10 over the multilingual long-document retrieval test sets; model records must preserve the reported dense, sparse, or hybrid configuration.

How this benchmark is scored

Categoryretrieval
Maximum score100 average nDCG@10
DirectionHigher is better

Primary source: https://arxiv.org/abs/2402.03216

Continue your research

Use these connected ModelRefs sections to compare alternatives, inspect implementation paths, and review the evidence and governance boundaries relevant to MLDR Leaderboard — AI Model Scores.