Benchmark

Translation Set1→en COMET22

language text 多语言

COMET-22 is a neural machine translation evaluation metric that uses an ensemble of two models: a COMET estimator trained with Direct Assessments and a multitask model that predicts sentence-level scores and word-level OK/BAD tags. It provides improved correlations with human judgments and increased robustness to critical errors compared to previous metrics.

语言EN
满分1
参评模型3

模型排名

名次 模型 机构 分数 来源
1 Nova Pro Amazon 89.0 来源 ↗
2 Nova Lite Amazon 88.8 来源 ↗
3 Nova Micro Amazon 88.7 来源 ↗