paper-with-me

Papers

Towards a Better Evaluation of Metrics for Machine Translation

2020-11-01 · WMT (EMNLP) 2020 11 · Peter Stanchev, Weiyue Wang, Hermann Ney

An important aspect of machine translation is its evaluation, which can be achieved through the use of a variety of metrics. To compare these metrics, the workshop on statistical machine translation annually evaluates metrics based on their correlation with human judgement. Over the years, methods for measuring correlation with humans have changed, but little research has been performed on what the optimal methods for acquiring human scores are and how human correlation can be measured. In this work, the methods for evaluating metrics at both system- and segment-level are analyzed in detail and their shortcomings are pointed out.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Machine TranslationTranslation

Similar Papers 제목 키워드 기반

Machine Translation Quality: A comparative evaluation of SMT, NMT and tailored-NMT outputs

2020-11-01 · EAMT 2020 11 · Maria Stasimioti, Vilelmini Sosoni, Katia Kermanidis, Despoina Mouratidis

The present study aims to compare three systems: a generic statistical machine translation (SMT), a generic neural machine translation (NMT) and a tailored-NMT system focusing on the English to Greek language pair. The c…

Machine TranslationNMTTranslation

Towards Explainable Evaluation Metrics for Machine Translation

2023-06-22 · Christoph Leiter, Piyawat Lertvittayakumjorn, Marina Fomicheva, Wei Zhao 외

Unlike classical lexical overlap metrics such as BLEU, most current evaluation metrics for machine translation (for example, COMET or BERTScore) are based on black-box large language models. They often achieve strong cor…

Machine TranslationTranslation

The Inside Story: Towards Better Understanding of Machine Translation Neural Evaluation Metrics

2023-05-19 · Ricardo Rei, Nuno M. Guerreiro, Marcos Treviso, Luisa Coheur 외

Neural metrics for machine translation evaluation, such as COMET, exhibit significant improvements in their correlation with human judgments, as compared to traditional metrics based on lexical overlap, such as BLEU. Yet…

Decision MakingMachine TranslationSentenceTranslation

Evaluating the Efficacy of Length-Controllable Machine Translation

2023-05-03 · Hao Cheng, Meng Zhang, Weixuan Wang, Liangyou Li 외

Length-controllable machine translation is a type of constrained translation. It aims to contain the original meaning as much as possible while controlling the length of the translation. We can use automatic summarizatio…

Machine TranslationTranslation

Adversarial Evaluation of Multimodal Machine Translation

2018-10-01 · EMNLP 2018 10 · Desmond Elliott

The promise of combining language and vision in multimodal machine translation is that systems will produce better translations by leveraging the image data. However, the evidence surrounding whether the images are usefu…

Machine TranslationMultimodal Machine Translationtext similarityTranslation