The AI arena is free today

Open Superagent

Translation Set1→en COMET22

Paper

Progress Over Time

Interactive timeline showing model performance evolution on Translation Set1→en COMET22

State-of-the-art frontier
Open
Proprietary

Translation Set1→en COMET22 Leaderboard

3 models
ContextCostLicense
1
Amazon
Amazon
2
Amazon
Amazon
3
Notice missing or incorrect data?
About this benchmark

What is Translation Set1→en COMET22?

COMET-22 is a neural machine translation evaluation metric that uses an ensemble of two models: a COMET estimator trained with Direct Assessments and a multitask model that predicts sentence-level scores and word-level OK/BAD tags. It provides improved correlations with human judgments and increased robustness to critical errors compared to previous metrics.

Translation Set1→en COMET22 is a text benchmark evaluating models on language tasks. LLM Stats tracks 3 models on this benchmark, scored on a 0–1 scale. The current average is 0.9, with the leader at 0.9.

Compare leaders on the best AI for language leaderboards.

Current leaders

Nova Pro from Amazon currently leads the Translation Set1→en COMET22 leaderboard with a score of 0.890 across 3 evaluated AI models.

1Nova ProAmazon89.0%
2Nova LiteAmazon88.8%
3Nova MicroAmazon88.7%

FAQ

Common questions about the Translation Set1→en COMET22 benchmark and leaderboard.

What is the Translation Set1→en COMET22 benchmark?

COMET-22 is a neural machine translation evaluation metric that uses an ensemble of two models: a COMET estimator trained with Direct Assessments and a multitask model that predicts sentence-level scores and word-level OK/BAD tags. It provides improved correlations with human judgments and increased robustness to critical errors compared to previous metrics.

What is the Translation Set1→en COMET22 leaderboard?

The Translation Set1→en COMET22 leaderboard ranks 3 AI models based on their performance on this benchmark. Currently, Nova Pro by Amazon leads with a score of 0.890. The average score across all models is 0.888.

What is the highest Translation Set1→en COMET22 score?

The highest Translation Set1→en COMET22 score is 0.890, achieved by Nova Pro from Amazon.

How many models are evaluated on Translation Set1→en COMET22?

3 models have been evaluated on the Translation Set1→en COMET22 benchmark, with 0 verified results and 3 self-reported results.

Where can I find the Translation Set1→en COMET22 paper?

The Translation Set1→en COMET22 paper is available at https://aclanthology.org/2022.wmt-1.52/. The paper details the methodology, dataset construction, and evaluation criteria.

What categories does Translation Set1→en COMET22 cover?

Translation Set1→en COMET22 is categorized under language. The benchmark evaluates text models with multilingual support.

How recent are the Translation Set1→en COMET22 leaderboard results?

The Translation Set1→en COMET22 leaderboard was last updated in August 2026 and currently includes 3 evaluated models.