English

An Automatic Evaluation of the WMT22 General Machine Translation Task

Computation and Language 2022-11-10 v2

Abstract

This report presents an automatic evaluation of the general machine translation task of the Seventh Conference on Machine Translation (WMT22). It evaluates a total of 185 systems for 21 translation directions including high-resource to low-resource language pairs and from closely related to distant languages. This large-scale automatic evaluation highlights some of the current limits of state-of-the-art machine translation systems. It also shows how automatic metrics, namely chrF, BLEU, and COMET, can complement themselves to mitigate their own limits in terms of interpretability and accuracy.

Keywords

Cite

@article{arxiv.2209.14172,
  title  = {An Automatic Evaluation of the WMT22 General Machine Translation Task},
  author = {Benjamin Marie},
  journal= {arXiv preprint arXiv:2209.14172},
  year   = {2022}
}

Comments

Update: correction, fr->de and de-> tables were switched

R2 v1 2026-06-28T02:17:52.951Z