paper-with-me

Papers

Evaluating Optimal Reference Translations

2023-11-28 · Vilém Zouhar, Věra Kloudová, Martin Popel, Ondřej Bojar

The overall translation quality reached by current machine translation (MT) systems for high-resourced language pairs is remarkably good. Standard methods of evaluation are not suitable nor intended to uncover the many translation errors and quality deficiencies that still persist. Furthermore, the quality of standard reference translations is commonly questioned and comparable quality levels have been reached by MT alone in several language pairs. Navigating further research in these high-resource settings is thus difficult. In this article, we propose a methodology for creating more reliable document-level human reference translations, called "optimal reference translations," with the simple aim to raise the bar of what should be deemed "human translation quality." We evaluate the obtained document-level optimal reference translations in comparison with "standard" ones, confirming a significant quality increase and also documenting the relationship between evaluation and translation editing.

📄 PDF Abstract BibTeX arXiv:2311.16787

Code (1)

ufal/optimal-reference-translations 공식 구현

Tasks

Machine TranslationTranslation

Similar Papers 제목 키워드 기반

Match without a Referee: Evaluating MT Adequacy without Reference Translations

2012-06-01 · WS 2012 6 · Yashar Mehdad, Matteo Negri, Marcello Federico
Machine TranslationNatural Language Inference

Manual and Automatic Paraphrases for MT Evaluation

2016-05-01 · LREC 2016 5 · Ale{\v{s}} Tamchyna, Petra Baran{\v{c}}{\'\i}kov{\'a}

Paraphrasing of reference translations has been shown to improve the correlation with human judgements in automatic evaluation of machine translation (MT) outputs. In this work, we present a new dataset for evaluating En…

Machine TranslationTranslation

On the Evaluation of Machine Translation n-best Lists

2020-11-01 · EMNLP (Eval4NLP) 2020 11 · Jacob Bremerman, Huda Khayrallah, Douglas Oard, Matt Post

The standard machine translation evaluation framework measures the single-best output of machine translation systems. There are, however, many situations where n-best lists are needed, yet there is no established way of …

Machine TranslationTranslationvalid

A Dataset for Probing Translationese Preferences in English-to-Swedish Translation

2026-03-09 · Jenny Kunz, Anja Jarochenko, Marcel Bollmann arxiv

Translations often carry traces of the source language, a phenomenon known as translationese. We introduce the first freely available English-to-Swedish dataset contrasting translationese sentences with idiomatic alterna…

Multi-Hypothesis Machine Translation Evaluation

2020-07-01 · ACL 2020 6 · Marina Fomicheva, Lucia Specia, Francisco Guzm{\'a}n

Reliably evaluating Machine Translation (MT) through automated metrics is a long-standing problem. One of the main challenges is the fact that multiple outputs can be equally valid. Attempts to minimise this issue includ…

Machine TranslationTranslationvalid