Integrating a Large, Monolingual Corpus as Translation Memory into Statistical Machine Translation
Code (0)
등록된 구현이 없습니다.
Tasks
Information RetrievalMachine TranslationTranslationSimilar Papers 제목 키워드 기반
Neural Machine Translation with Monolingual Translation Memory
Prior work has proved that Translation memory (TM) can boost the performance of Neural Machine Translation (NMT). In contrast to existing work that uses bilingual corpus as TM and employs source-side similarity search fo…
Domain AdaptationMachine TranslationNMTRetrieval+1High-Quality Data Augmentation for Low-Resource NMT: Combining a Translation Memory, a GAN Generator, and Filtering
Back translation, as a technique for extending a dataset, is widely used by researchers in low-resource language translation tasks. It typically translates from the target to the source language to ensure high-quality tr…
Data AugmentationGenerative Adversarial NetworkLow Resource NMTMachine Translation+3Fully Synthetic Data Improves Neural Machine Translation with Knowledge Distillation
This paper explores augmenting monolingual data for knowledge distillation in neural machine translation. Source language monolingual text can be incorporated as a forward translation. Interestingly, we find the best way…
Knowledge DistillationMachine TranslationTranslationArzEn-ST: A Three-way Speech Translation Corpus for Code-Switched Egyptian Arabic - English
We present our work on collecting ArzEn-ST, a code-switched Egyptian Arabic - English Speech Translation Corpus. This corpus is an extension of the ArzEn speech corpus, which was collected through informal interviews wit…
Machine TranslationTranslationEnhancement of Encoder and Attention Using Target Monolingual Corpora in Neural Machine Translation
A large-scale parallel corpus is required to train encoder-decoder neural machine translation. The method of using synthetic parallel texts, in which target monolingual corpora are automatically translated into source se…
DecoderDiversityLanguage ModelingLanguage Modelling+3