all the models — AI benchmark observatory
← Benchmarks

Translation en→Set1 COMET22

COMET-22 is an ensemble machine translation evaluation metric combining a COMET estimator model trained with Direct Assessments and a multitask model that predicts sentence-level scores and word-level OK/BAD tags. It demonstrates improved correlations compared to state-of-the-art metrics and increased robustness to critical errors.

id translation-en→set1-comet22 · max 1 · 3 models reported

#ModelScore

No scores for this benchmark yet.

Translation en→Set1 COMET22 Leaderboard · all the models