paper-with-me

Papers

Integrating a Large, Monolingual Corpus as Translation Memory into Statistical Machine Translation

2015-05-01 · WS 2015 5 · Katharina W{\"a}schle, Stefan Riezler
📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Information RetrievalMachine TranslationTranslation

Similar Papers 제목 키워드 기반

Neural Machine Translation with Monolingual Translation Memory

2021-05-24 · ACL 2021 5 · Deng Cai, Yan Wang, Huayang Li, Wai Lam 외

Prior work has proved that Translation memory (TM) can boost the performance of Neural Machine Translation (NMT). In contrast to existing work that uses bilingual corpus as TM and employs source-side similarity search fo…

Domain AdaptationMachine TranslationNMTRetrieval+1

High-Quality Data Augmentation for Low-Resource NMT: Combining a Translation Memory, a GAN Generator, and Filtering

2024-08-22 · Hengjie Liu, Ruibo Hou, Yves Lepage

Back translation, as a technique for extending a dataset, is widely used by researchers in low-resource language translation tasks. It typically translates from the target to the source language to ensure high-quality tr…

Data AugmentationGenerative Adversarial NetworkLow Resource NMTMachine Translation+3

Fully Synthetic Data Improves Neural Machine Translation with Knowledge Distillation

2020-12-31 · Alham Fikri Aji, Kenneth Heafield

This paper explores augmenting monolingual data for knowledge distillation in neural machine translation. Source language monolingual text can be incorporated as a forward translation. Interestingly, we find the best way…

Knowledge DistillationMachine TranslationTranslation

ArzEn-ST: A Three-way Speech Translation Corpus for Code-Switched Egyptian Arabic - English

2022-11-22 · Injy Hamed, Nizar Habash, Slim Abdennadher, Ngoc Thang Vu

We present our work on collecting ArzEn-ST, a code-switched Egyptian Arabic - English Speech Translation Corpus. This corpus is an extension of the ArzEn speech corpus, which was collected through informal interviews wit…

Machine TranslationTranslation

Enhancement of Encoder and Attention Using Target Monolingual Corpora in Neural Machine Translation

2018-07-01 · WS 2018 7 · Kenji Imamura, Atsushi Fujita, Eiichiro Sumita

A large-scale parallel corpus is required to train encoder-decoder neural machine translation. The method of using synthetic parallel texts, in which target monolingual corpora are automatically translated into source se…

DecoderDiversityLanguage ModelingLanguage Modelling+3