Prompsit’s Submission to the IWSLT 2018 Low Resource Machine Translation Task
This paper presents Prompsit Language Engineering’s submission to the IWSLT 2018 Low Resource Machine Translation task. Our submission is based on cross-lingual learning: a multilingual neural machine translation system was created with the sole purpose of improving translation quality on the Basque-to-English language pair. The multilingual system was trained on a combination of in-domain data, pseudo in-domain data obtained via cross-entropy data selection and backtranslated data. We morphologically segmented Basque text with a novel approach that only requires a dictionary such as those used by spell checkers and proved that this segmentation approach outperforms the widespread byte pair encoding strategy for this task.
Code (0)
등록된 구현이 없습니다.
Tasks
Machine TranslationTranslationSimilar Papers 제목 키워드 기반
Prompsit's submission to WMT 2018 Parallel Corpus Filtering shared task
This paper describes Prompsit Language Engineering{'}s submissions to the WMT 2018 parallel corpus filtering shared task. Our four submissions were based on an automatic classifier for identifying pairs of sentences that…
Active LearningLanguage ModelingLanguage ModellingMachine TranslationNAVER LABS Europe's Multilingual Speech Translation Systems for the IWSLT 2023 Low-Resource Track
This paper presents NAVER LABS Europe's systems for Tamasheq-French and Quechua-Spanish speech translation in the IWSLT 2023 Low-Resource track. Our work attempts to maximize translation quality in low-resource settings …
TranslationCUNI Basque-to-English Submission in IWSLT18
We present our submission to the IWSLT18 Low Resource task focused on the translation from Basque-to-English. Our submission is based on the current state-of-the-art self-attentive neural network architecture, Transforme…
Transfer LearningTranslationIMS' Systems for the IWSLT 2021 Low-Resource Speech Translation Task
This paper describes the submission to the IWSLT 2021 Low-Resource Speech Translation Shared Task by IMS team. We utilize state-of-the-art models combined with several data augmentation, multi-task and transfer learning …
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Data AugmentationMachine Translation+4IMS’ Systems for the IWSLT 2021 Low-Resource Speech Translation Task
This paper describes the submission to the IWSLT 2021 Low-Resource Speech Translation Shared Task by IMS team. We utilize state-of-the-art models combined with several data augmentation, multi-task and transfer learning …
Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Data AugmentationMachine Translation+4