HW-TSC’s Participation in the WMT 2021 Large-Scale Multilingual Translation Task
This paper presents the submission of Huawei Translation Services Center (HW-TSC) to the WMT 2021 Large-Scale Multilingual Translation Task. We participate in Samll Track #2, including 6 languages: Javanese (Jv), Indonesian (Id), Malay (Ms), Tagalog (Tl), Tamil (Ta) and English (En) with 30 directions under the constrained condition. We use Transformer architecture and obtain the best performance via multiple variants with larger parameter sizes. We train a single multilingual model to translate all the 30 directions. We perform detailed pre-processing and filtering on the provided large-scale bilingual and monolingual datasets. Several commonly used strategies are used to train our models, such as Back Translation, Forward Translation, Ensemble Knowledge Distillation, Adapter Fine-tuning. Our model obtains competitive results in the end.
Code (0)
등록된 구현이 없습니다.
Tasks
Knowledge DistillationTranslationSimilar Papers 제목 키워드 기반
Rakuten’s Participation in WAT 2021: Examining the Effectiveness of Pre-trained Models for Multilingual and Multimodal Machine Translation
This paper introduces our neural machine translation systems’ participation in the WAT 2021 shared translation tasks (team ID: sakura). We participated in the (i) NICT-SAP, (ii) Japanese-English multimodal translation, (…
DenoisingLanguage ModelingLanguage ModellingMachine Translation+2HW-TSC’s Participation in the WMT 2021 Triangular MT Shared Task
This paper presents the submission of Huawei Translation Service Center (HW-TSC) to WMT 2021 Triangular MT Shared Task. We participate in the Russian-to-Chinese task under the constrained condition. We use Transformer ar…
DenoisingTranslationHW-TSC’s Participation in the WMT 2021 News Translation Shared Task
This paper presents the submission of Huawei Translate Services Center (HW-TSC) to the WMT 2021 News Translation Shared Task. We participate in 7 language pairs, including Zh/En, De/En, Ja/En, Ha/En, Is/En, Hi/Bn, and Xh…
de-enKnowledge DistillationTranslationNICT's participation to WAT 2019: Multilingualism and Multi-step Fine-Tuning for Low Resource NMT
In this paper we describe our submissions to WAT 2019 for the following tasks: English{--}Tamil translation and Russian{--}Japanese translation. Our team,{``}NICT-5{''}, focused on multilingual domain adaptation and back…
Domain AdaptationLow Resource NMTNMTTranslationThe TALP-UPC System Description for WMT20 News Translation Task: Multilingual Adaptation for Low Resource MT
In this article, we describe the TALP-UPC participation in the WMT20 news translation shared task for Tamil-English. Given the low amount of parallel training data, we resort to adapt the task to a multilingual system to…
Translation