A Document-Level Neural Machine Translation Model with Dynamic Caching Guided by Theme-Rheme Information
Research on document-level Neural Machine Translation (NMT) models has attracted increasing attention in recent years. Although the proposed works have proved that the inter-sentence information is helpful for improving the performance of the NMT models, what information should be regarded as context remains ambiguous. To solve this problem, we proposed a novel cache-based document-level NMT model which conducts dynamic caching guided by theme-rheme information. The experiments on NIST evaluation sets demonstrate that our proposed model achieves substantial improvements over the state-of-the-art baseline NMT models. As far as we know, we are the first to introduce theme-rheme theory into the field of machine translation.
Code (1)
Tasks
Machine TranslationNMTSentenceTranslationSimilar Papers 제목 키워드 기반
Dynamic Context Selection for Document-level Neural Machine Translation via Reinforcement Learning
Document-level neural machine translation has yielded attractive improvements. However, majority of existing methods roughly use all context sentences in a fixed scope. They neglect the fact that different source sentenc…
Machine Translationreinforcement-learningReinforcement Learning (RL)Sentence+1STAR : Sentence Translation Alignment Rate for Document-to-Document Machine Translation
Large Language Models (LLMs) have enabled a shift from sentence-level to document-to-document (Doc2Doc) machine translation, promising improved global coherence. However, document-to-document generation in a single pass …
Machine TranslationLog-Linear Reformulation of the Noisy Channel Model for Document-Level Neural Machine Translation
We seek to maximally use various data sources, such as parallel and monolingual data, to build an effective and efficient document-level translation system. In particular, we start by considering a noisy channel approach…
Language ModelingLanguage ModellingMachine TranslationSentence+1Document-Level Neural Machine Translation Using BERT as Context Encoder
Large-scale pre-trained representations such as BERT have been widely used in many natural language understanding tasks. The methods of incorporating BERT into document-level machine translation are still being explored.…
DecoderDocument Level Machine TranslationMachine TranslationNatural Language Understanding+2Multilingual Contextualization of Large Language Models for Document-Level Machine Translation
Large language models (LLMs) have demonstrated strong performance in sentence-level machine translation, but scaling to document-level translation remains challenging, particularly in modeling long-range dependencies and…
Document Level Machine TranslationDocument TranslationMachine TranslationSentence+1