paper-with-me

홈 › Papers

Code-Switching for Enhancing NMT with Pre-Specified Translation

2019-04-19 · NAACL 2019 6 · Kai Song, Yue Zhang, Heng Yu, Weihua Luo, Kun Wang, Min Zhang

Leveraging user-provided translation to constrain NMT has practical significance. Existing methods can be classified into two main categories, namely the use of placeholder tags for lexicon words and the use of hard constraints during decoding. Both methods can hurt translation fidelity for various reasons. We investigate a data augmentation method, making code-switched training data by replacing source phrases with their target translations. Our method does not change the MNT model or decoding algorithm, allowing the model to learn lexicon translations by copying source-side target words. Extensive experiments show that our method achieves consistent improvements over existing approaches, improving translation of constrained words without hurting unconstrained words.

📄 PDF Abstract BibTeX arXiv:1904.09107

Code (1)

batman2013/e-commerce_test_sets 공식 구현

Tasks

Data AugmentationNMTTranslation

Similar Papers 제목 키워드 기반

CoVoSwitch: Machine Translation of Synthetic Code-Switched Text Based on Intonation Units

2024-07-19 · Yeeun Kang

Multilingual code-switching research is often hindered by the lack and linguistically biased status of available datasets. To expand language representation, we synthesize code-switching data by replacing intonation unit…

Machine TranslationSpeech-to-TextSpeech-to-Text TranslationTranslation

Multi-Domain Neural Machine Translation

2018-05-06 · Sander Tars, Mark Fishel

We present an approach to neural machine translation (NMT) that supports multiple domains in a single model and allows switching between the domains when translating. The core idea is to treat text domains as distinct la…

Machine TranslationNMTTranslation

Code-Switching without Switching: Language Agnostic End-to-End Speech Translation

2022-10-04 · Christian Huber, Enes Yavuz Ugan, Alexander Waibel

We propose a) a Language Agnostic end-to-end Speech Translation model (LAST), and b) a data augmentation strategy to increase code-switching (CS) performance. With increasing globalization, multiple languages are increas…

Data Augmentationspeech-recognitionSpeech RecognitionTranslation

Low-resource Machine Translation for Code-switched Kazakh-Russian Language Pair

2025-03-25 · Maksim Borisov, Zhanibek Kozhirbayev, Valentin Malykh

Machine translation for low resource language pairs is a challenging task. This task could become extremely difficult once a speaker uses code switching. We propose a method to build a machine translation model for code-…

Machine TranslationTranslation

Contextual Code Switching for Machine Translation using Language Models

2023-12-20 · Arshad Kaji, Manan Shah

Large language models (LLMs) have exerted a considerable impact on diverse language-related tasks in recent years. Their demonstrated state-of-the-art performance is achieved through methodologies such as zero-shot or fe…

Machine TranslationQuestion AnsweringTranslation