all the models — AI benchmark observatory
← Benchmarks

Global-MMLU

A comprehensive multilingual benchmark covering 42 languages that addresses cultural and linguistic biases in evaluation, with improved translation quality and culturally sensitive question subsets.

id global-mmlu · max 1 · 6 models reported

#ModelScore
1Claude Opus 5.5
Anthropic
0.94