← Benchmarks
Translation Set1→en COMET22
COMET-22 is a neural machine translation evaluation metric that uses an ensemble of two models: a COMET estimator trained with Direct Assessments and a multitask model that predicts sentence-level scores and word-level OK/BAD tags. It provides improved correlations with human judgments and increased robustness to critical errors compared to previous metrics.
id translation-set1→en-comet22 · max 1 · 3 models reported
| # | Model | Score |
|---|
No scores for this benchmark yet.
