all the models — AI benchmark observatory
← Benchmarks

MMLU-ProX

Extended version of MMLU-Pro providing additional challenging multiple-choice questions for evaluating language models across diverse academic and professional domains. Built on the foundation of the Massive Multitask Language Understanding benchmark framework.

id mmlu-prox · max 1 · 32 models reported

#ModelScore

No scores for this benchmark yet.

MMLU-ProX Leaderboard · all the models