all the models — AI benchmark observatory
← Benchmarks

Multilingual MMLU

MMLU-ProX is a comprehensive multilingual benchmark covering 29 typologically diverse languages, building upon MMLU-Pro. Each language version consists of 11,829 identical questions enabling direct cross-linguistic comparisons. The benchmark evaluates large language models' reasoning capabilities across linguistic and cultural boundaries through challenging, reasoning-focused questions with 10 answer choices.

id multilingual-mmlu · max 1 · 5 models reported

#ModelScore

No scores for this benchmark yet.

Multilingual MMLU Leaderboard · all the models