paper-with-me

홈 › Papers

Machine Translation Human Evaluation: an investigation of evaluation based on Post-Editing and its relation with Direct Assessment

2018-10-01 · IWSLT (EMNLP) 2018 10 · Luisa Bentivogli, Mauro Cettolo, Marcello Federico, Christian Federmann

In this paper we present an analysis of the two most prominent methodologies used for the human evaluation of MT quality, namely evaluation based on Post-Editing (PE) and evaluation based on Direct Assessment (DA). To this purpose, we exploit a publicly available large dataset containing both types of evaluations. We first focus on PE and investigate how sensitive TER-based evaluation is to the type and number of references used. Then, we carry out a comparative analysis of PE and DA to investigate the extent to which the evaluation results obtained by methodologies addressing different human perspectives are similar. This comparison sheds light not only on PE but also on the so-called reference bias related to monolingual DA. Also, we analyze if and how the two methodologies can complement each other’s weaknesses.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Machine Translation

Similar Papers 제목 키워드 기반

A Set of Recommendations for Assessing Human-Machine Parity in Language Translation

2020-04-03 · Samuel Läubli, Sheila Castilho, Graham Neubig, Rico Sennrich 외

The quality of machine translation has increased remarkably over the past years, to the degree that it was found to be indistinguishable from professional human translation in a number of empirical investigations. We rea…

Machine TranslationTranslation

Further Investigation into Reference Bias in Monolingual Evaluation of Machine Translation

2017-09-01 · EMNLP 2017 9 · Qingsong Ma, Yvette Graham, Timothy Baldwin, Qun Liu

Monolingual evaluation of Machine Translation (MT) aims to simplify human assessment by requiring assessors to compare the meaning of the MT output with a reference translation, opening up the task to a much larger pool …

Machine TranslationTranslation

Statistical Power and Translationese in Machine Translation Evaluation

2020-11-01 · EMNLP 2020 11 · Yvette Graham, Barry Haddow, Philipp Koehn

The term translationese has been used to describe features of translated text, and in this paper, we provide detailed analysis of potential adverse effects of translationese on machine translation evaluation. Our analysi…

Machine TranslationTranslation

Evaluating the IWSLT2023 Speech Translation Tasks: Human Annotations, Automatic Metrics, and Segmentation

2024-06-06 · Matthias Sperber, Ondřej Bojar, Barry Haddow, Dávid Javorský 외

Human evaluation is a critical component in machine translation system development and has received much attention in text translation research. However, little prior work exists on the topic of human evaluation for spee…

Machine TranslationTranslation

Beyond Human-Only: Evaluating Human-Machine Collaboration for Collecting High-Quality Translation Data

2024-10-14 · Zhongtao Liu, Parker Riley, Daniel Deutsch, Alison Lui 외

Collecting high-quality translations is crucial for the development and evaluation of machine translation systems. However, traditional human-only approaches are costly and slow. This study presents a comprehensive inves…

Machine TranslationTranslation