all the models — AI benchmark observatory
← Benchmarks

WMT24++

WMT24++ is a comprehensive multilingual machine translation benchmark that expands the WMT24 dataset to cover 55 languages and dialects. It includes human-written references and post-edits across four domains (literary, news, social, and speech) to evaluate machine translation systems and large language models across diverse linguistic contexts.

id wmt24++ · max 1 · 24 models reported

#ModelScore

No scores for this benchmark yet.

WMT24++ Leaderboard · all the models