all the models — AI benchmark observatory
← Benchmarks

WMT23

The Eighth Conference on Machine Translation (WMT23) benchmark evaluating machine translation systems across 8 language pairs (14 translation directions) including general, biomedical, literary, and low-resource language translation tasks. Features specialized shared tasks for quality estimation, metrics evaluation, sign language translation, and discourse-level literary translation with professional human assessment.

id wmt23 · max 1 · 4 models reported

#ModelScore

No scores for this benchmark yet.