Exploring Adequacy Errors in Neural Machine Translation with the Help of Cross-Language Aligned Word Embeddings
Neural machine translation (NMT) was shown to produce more fluent output than phrase-based statistical (PBMT) and rule-based machine translation (RBMT). However, improved fluency makes it more difficult for post editors to identify and correct adequacy errors, because unlike RBMT and SMT, in NMT adequacy errors are frequently not anticipated by fluency errors. Omissions and additions of content in otherwise flawlessly fluent NMT output are the most prominent types of such adequacy errors, which can only be detected with reference to source texts. This contribution explores the degree of semantic similarity between source texts, NMT output and post edited output. In this way, computational semantic similarity scores (cosine similarity) are related to human quality judgments. The analyses are based on publicly available NMT post editing data annotated for errors in three language pairs (EN-DE, EN-LV, EN-HR) with the Multidimensional Quality Metrics (MQM). Methodologically, this contribution tests whether cross-language aligned word embeddings as the sole source of semantic information mirror human error annotation.
Code (0)
등록된 구현이 없습니다.
Tasks
de-enMachine TranslationNMTSemantic SimilaritySemantic Textual SimilarityTranslationWord EmbeddingsSimilar Papers 제목 키워드 기반
Relations between comprehensibility and adequacy errors in machine translation output
This work presents a detailed analysis of translation errors perceived by readers as comprehensibility and/or adequacy issues. The main finding is that good comprehensibility, similarly to good fluency, can mask a number…
Machine TranslationTranslationWord TranslationBLEU Evaluation of Machine-Translated English-Croatian Legislation
This paper presents work on the evaluation of online available machine translation (MT) service, i.e. Google Translate, for English-Croatian language pair in the domain of legislation. The total set of 200 sentences, for…
Machine TranslationTranslationDetecting over/under-translation errors for determining adequacy in human translations
We present a novel approach to detecting over and under translations (OT/UT) as part of adequacy error checks in translation evaluation. We do not restrict ourselves to machine translation (MT) outputs and specifically t…
Language ModelingLanguage ModellingMachine TranslationTranslationExploring the Importance of Source Text in Automatic Post-Editing for Context-Aware Machine Translation
Accurate translation requires document-level information, which is ignored by sentence-level machine translation. Recent work has demonstrated that document-level consistency can be improved with automatic post-editing (…
Automatic Post-EditingMachine TranslationSentenceTranslationRevisit Automatic Error Detection for Wrong and Missing Translation -- A Supervised Approach
While achieving great fluency, current machine translation (MT) techniques are bottle-necked by adequacy issues. To have a closer study of these issues and accelerate model development, we propose automatic detecting ade…
Machine TranslationTranslation