Benchmark

Translation en→Set1 COMET22

COMET-22 is an ensemble machine translation evaluation metric combining a COMET estimator model trained with Direct Assessments and a multitask model that…

Modality
text
Categories
language
Openness
unknown
Source
llm_stats
Reported scores
3

Reported scores

llm_stats

ModelOrganizationReported valueReportedEvidence
Nova LiteAmazon0.8882024-11-20evidence
Nova MicroAmazon0.8852024-11-20evidence
Nova ProAmazon0.8912024-11-20evidence

Scores are partitioned by the source that reported them and are never merged into a single cross-source ranking, because the sources measure different things and say so.