Quality Estimation and Translation Metrics via Pre-trained Word and Sentence Embeddings
We propose the use of pre-trained embeddings as features of a regression model for sentence-level quality estimation of machine translation. In our work we combine freely available BERT and LASER multilingual embeddings to train a neural-based regression model. In the second proposed method we use as an input features not only pre-trained embeddings, but also log probability of any machine translation (MT) system. Both methods are applied to several language pairs and are evaluated both as a classical quality estimation system (predicting the HTER score) as well as an MT metric (predicting human judgements of translation quality).
Code (0)
등록된 구현이 없습니다.
Tasks
Machine TranslationregressionSentenceSentence EmbeddingsTranslationMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Metrics for Evaluation of Word-level Machine Translation Quality Estimation
Findings of the WMT 2019 Shared Tasks on Quality Estimation
We report the results of the WMT19 shared task on Quality Estimation, i.e. the task of predicting the quality of the output of machine translation systems given just the source text and the hypothesis translations. The t…
Machine TranslationSentenceTranslationRevisiting Round-Trip Translation for Quality Estimation
Quality estimation (QE) is the task of automatically evaluating the quality of translations without human-translated references. Calculating BLEU between the input sentence and round-trip translation (RTT) was once consi…
NMTSentenceSentence EmbeddingsTranslationUnsupervised Word-level Quality Estimation for Machine Translation Through the Lens of Annotators (Dis)agreement
Word-level quality estimation (WQE) aims to automatically identify fine-grained error spans in machine-translated outputs and has found many uses, including assisting translators during post-editing. Modern WQE technique…
Language ModelingLanguage ModellingMachine TranslationTranslation+1SentSim: Crosslingual Semantic Evaluation of Machine Translation
Machine translation (MT) is currently evaluated in one of two ways: in a monolingual fashion, by comparison with the system output to one or more human reference translations, or in a trained crosslingual fashion, by bui…
Machine TranslationSemantic SimilaritySemantic Textual SimilaritySentence+1