paper-with-me

홈 › Papers

Quality and Quantity of Machine Translation References for Automatic Metrics

2024-01-02 · Vilém Zouhar, Ondřej Bojar

Automatic machine translation metrics typically rely on human translations to determine the quality of system translations. Common wisdom in the field dictates that the human references should be of very high quality. However, there are no cost-benefit analyses that could be used to guide practitioners who plan to collect references for machine translation evaluation. We find that higher-quality references lead to better metric correlations with humans at the segment-level. Having up to 7 references per segment and taking their average (or maximum) helps all metrics. Interestingly, the references from vendors of different qualities can be mixed together and improve metric success. Higher quality references, however, cost more to create and we frame this as an optimization problem: given a specific budget, what references should be collected to maximize metric success. These findings can be used by evaluators of shared tasks when references need to be created under a certain budget.

📄 PDF Abstract BibTeX arXiv:2401.01283

Code (1)

ufal/optimal-reference-translations 공식 구현

Tasks

Machine TranslationTranslation

Similar Papers 제목 키워드 기반

Optimizing Statistical Machine Translation for Text Simplification

2016-01-01 · TACL 2016 1 · Wei Xu, Courtney Napoles, Ellie Pavlick, Quanze Chen 외

Most recent sentence simplification systems use basic machine translation models to learn lexical and syntactic paraphrases from a manually simplified parallel corpus. These methods are limited by the quality and quantit…

Machine TranslationSentenceText SimplificationTranslation

Modeling User Preferences with Automatic Metrics: Creating a High-Quality Preference Dataset for Machine Translation

2024-10-10 · Sweta Agrawal, José G. C. de Souza, Ricardo Rei, António Farinhas 외

Alignment with human preferences is an important step in developing accurate and safe large language models. This is no exception in machine translation (MT), where better handling of language nuances and context-specifi…

Machine TranslationSentenceTranslation

A Quality-based Active Sample Selection Strategy for Statistical Machine Translation

2014-05-01 · LREC 2014 5 · Varvara Logacheva, Lucia Specia

This paper presents a new active learning technique for machine translation based on quality estimation of automatically translated sentences. It uses an error-driven strategy, i.e., it assumes that the more errors an au…

Active LearningMachine TranslationSentenceSentiment Analysis+1

A Post-Editing Dataset in the Legal Domain: Do we Underestimate Neural Machine Translation Quality?

2020-05-01 · LREC 2020 5 · Julia Ive, Lucia Specia, Sara Szoc, Tom Vanallemeersch 외

We introduce a machine translation dataset for three pairs of languages in the legal domain with post-edited high-quality neural machine translation and independent human references. The data was collected as part of the…

Automatic Post-EditingMachine TranslationSentenceTranslation

BLEU might be Guilty but References are not Innocent

2020-04-13 · EMNLP 2020 11 · Markus Freitag, David Grangier, Isaac Caswell

The quality of automatic metrics for machine translation has been increasingly called into question, especially for high-quality systems. This paper demonstrates that, while choice of metric is important, the nature of t…

DiversityMachine TranslationTranslation