ModelRefs / FLORES-200 Leaderboard — AI Model Scores

FLORES-200 Leaderboard — AI Model Scores

Machine translation eval across 200 languages. Current leaders, methodology, and citation sources for FLORES-200.

Overview

Machine translation eval across 200 languages.

How it is measured: spBLEU mean across all language pairs.

How this benchmark is scored

Categoryreasoning
Maximum score100 spBLEU
DirectionHigher is better

Primary source: https://github.com/facebookresearch/flores

Continue your research

Use these connected ModelRefs sections to compare alternatives, inspect implementation paths, and review the evidence and governance boundaries relevant to FLORES-200 Leaderboard — AI Model Scores.