paper-with-me

홈 › Papers

Exploring Adequacy Errors in Neural Machine Translation with the Help of Cross-Language Aligned Word Embeddings

2019-09-01 · RANLP 2019 9 · Michael Ustaszewski

Neural machine translation (NMT) was shown to produce more fluent output than phrase-based statistical (PBMT) and rule-based machine translation (RBMT). However, improved fluency makes it more difficult for post editors to identify and correct adequacy errors, because unlike RBMT and SMT, in NMT adequacy errors are frequently not anticipated by fluency errors. Omissions and additions of content in otherwise flawlessly fluent NMT output are the most prominent types of such adequacy errors, which can only be detected with reference to source texts. This contribution explores the degree of semantic similarity between source texts, NMT output and post edited output. In this way, computational semantic similarity scores (cosine similarity) are related to human quality judgments. The analyses are based on publicly available NMT post editing data annotated for errors in three language pairs (EN-DE, EN-LV, EN-HR) with the Multidimensional Quality Metrics (MQM). Methodologically, this contribution tests whether cross-language aligned word embeddings as the sole source of semantic information mirror human error annotation.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

de-enMachine TranslationNMTSemantic SimilaritySemantic Textual SimilarityTranslationWord Embeddings

Similar Papers 제목 키워드 기반

Relations between comprehensibility and adequacy errors in machine translation output

2020-11-01 · CONLL 2020 · Maja Popovi{\'c}

This work presents a detailed analysis of translation errors perceived by readers as comprehensibility and/or adequacy issues. The main finding is that good comprehensibility, similarly to good fluency, can mask a number…

Machine TranslationTranslationWord Translation

BLEU Evaluation of Machine-Translated English-Croatian Legislation

2012-05-01 · LREC 2012 5 · Sanja Seljan, Marija Brki{\'c}, Tomislav Vi{\v{c}}i{\'c}

This paper presents work on the evaluation of online available machine translation (MT) service, i.e. Google Translate, for English-Croatian language pair in the domain of legislation. The total set of 200 sentences, for…

Machine TranslationTranslation

Detecting over/under-translation errors for determining adequacy in human translations

2021-04-01 · Prabhakar Gupta, Ridha Juneja, Anil Nelakanti, Tamojit Chatterjee

We present a novel approach to detecting over and under translations (OT/UT) as part of adequacy error checks in translation evaluation. We do not restrict ourselves to machine translation (MT) outputs and specifically t…

Language ModelingLanguage ModellingMachine TranslationTranslation

Exploring the Importance of Source Text in Automatic Post-Editing for Context-Aware Machine Translation

2021-05-01 · NoDaLiDa 2021 5 · Chaojun Wang, Christian Hardmeier, Rico Sennrich

Accurate translation requires document-level information, which is ignored by sentence-level machine translation. Recent work has demonstrated that document-level consistency can be improved with automatic post-editing (…

Automatic Post-EditingMachine TranslationSentenceTranslation

Revisit Automatic Error Detection for Wrong and Missing Translation -- A Supervised Approach

2019-11-01 · IJCNLP 2019 11 · Wenqiang Lei, Weiwen Xu, Ai Ti Aw, Yuanxin Xiang 외

While achieving great fluency, current machine translation (MT) techniques are bottle-necked by adequacy issues. To have a closer study of these issues and accelerate model development, we propose automatic detecting ade…

Machine TranslationTranslation