← Benchmarks
MMLU-ProX
Extended version of MMLU-Pro providing additional challenging multiple-choice questions for evaluating language models across diverse academic and professional domains. Built on the foundation of the Massive Multitask Language Understanding benchmark framework.
id mmlu-prox · max 1 · 32 models reported
| # | Model | Score |
|---|
No scores for this benchmark yet.
