paper-with-me

Papers

Adaptively Scheduled Multitask Learning: The Case of Low-Resource Neural Machine Translation

2019-11-01 · WS 2019 11 · Poorya Zaremoodi, Gholamreza Haffari

Neural Machine Translation (NMT), a data-hungry technology, suffers from the lack of bilingual data in low-resource scenarios. Multitask learning (MTL) can alleviate this issue by injecting inductive biases into NMT, using auxiliary syntactic and semantic tasks. However, an effective \textit{training schedule} is required to balance the importance of tasks to get the best use of the training signal. The role of training schedule becomes even more crucial in \textit{biased-MTL} where the goal is to improve one (or a subset) of tasks the most, e.g. translation quality. Current approaches for biased-MTL are based on brittle \textit{hand-engineered} heuristics that require trial and error, and should be (re-)designed for each learning scenario. To the best of our knowledge, ours is the first work on \textit{adaptively} and \textit{dynamically} changing the training schedule in biased-MTL. We propose a rigorous approach for automatically reweighing the training data of the main and auxiliary tasks throughout the training process based on their contributions to the generalisability of the main NMT task. Our experiments on translating from English to Vietnamese/Turkish/Spanish show improvements of up to +1.2 BLEU points, compared to strong baselines. Additionally, our analyses shed light on the dynamic of needs throughout the training of NMT: from syntax to semantic.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Low Resource Neural Machine TranslationLow-Resource Neural Machine TranslationMachine TranslationNMTTranslation

Similar Papers 제목 키워드 기반

Risk-Aware Resource Allocation for URLLC: Challenges and Strategies with Machine Learning

2018-12-22 · Amin Azari, Mustafa Ozger, Cicek Cavdar

Supporting ultra-reliable low-latency communications (URLLC) is a major challenge of 5G wireless networks. Stringent delay and reliability requirements need to be satisfied for both scheduled and non-scheduled URLLC traf…

BIG-bench Machine LearningManagement

Cross-Modal Subspace Learning with Scheduled Adaptive Margin Constraints

2019-09-30 · David Semedo, João Magalhães

Cross-modal embeddings, between textual and visual modalities, aim to organise multimodal instances by their semantic correlations. State-of-the-art approaches use maximum-margin methods, based on the hinge-loss, to enfo…

Incremental LearningTriplet

MultiTask-CenterNet (MCN): Efficient and Diverse Multitask Learning using an Anchor Free Approach

2021-08-11 · Falk Heuer, Sven Mantowsky, Syed Saqib Bukhari, Georg Schneider

Multitask learning is a common approach in machine learning, which allows to train multiple objectives with a shared architecture. It has been shown that by training multiple tasks together inference time and compute res…

Depth Estimationobject-detectionObject DetectionPose Estimation+1

Challenges in Developing LRs for Non-Scheduled Languages: A Case of Magahi

2021-11-30 · Ritesh Kumar

Magahi is an Indo-Aryan Language, spoken mainly in the Eastern parts of India. Despite having a significant number of speakers, there has been virtually no language resource (LR) or language technology (LT) developed for…

POS

MLM: A Benchmark Dataset for Multitask Learning with Multiple Languages and Modalities

2020-08-14 · Jason Armitage, Endri Kacupaj, Golsa Tahmasebzadeh, Swati 외

In this paper, we introduce the MLM (Multiple Languages and Modalities) dataset - a new resource to train and evaluate multitask systems on samples in multiple modalities and three languages. The generation process and i…

Representation LearningScene Understanding