paper-with-me

홈 › Papers

Being Generous with Sub-Words towards Small NMT Children

2020-05-01 · LREC 2020 5 · Arne Defauw, Tom Vanallemeersch, Koen Van Winckel, Sara Szoc, Joachim Van den Bogaert

In the context of under-resourced neural machine translation (NMT), transfer learning from an NMT model trained on a high resource language pair, or from a multilingual NMT (M-NMT) model, has been shown to boost performance to a large extent. In this paper, we focus on so-called cold start transfer learning from an M-NMT model, which means that the parent model is not trained on any of the child data. Such a set-up enables quick adaptation of M-NMT models to new languages. We investigate the effectiveness of cold start transfer learning from a many-to-many M-NMT model to an under-resourced child. We show that sufficiently large sub-word vocabularies should be used for transfer learning to be effective in such a scenario. When adopting relatively large sub-word vocabularies we observe increases in performance thanks to transfer learning from a parent M-NMT model, both when translating to and from the under-resourced language. Our proposed approach involving dynamic vocabularies is both practical and effective. We report results on two under-resourced language pairs, i.e. Icelandic-English and Irish-English.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Machine TranslationNMTTransfer LearningTranslation

Similar Papers 제목 키워드 기반

Baby Scale: Investigating Models Trained on Individual Children's Language Input

2026-03-31 · Steven Y. Feng, Alvin W. M. Tan, Michael C. Frank arxiv

Modern language models (LMs) must be trained on many orders of magnitude more words of training data than human children receive before they begin to produce useful behavior. Assessing the nature and origins of this "dat…

Language Acquisition

Welfare Reform: Consequences for the Children

2025-06-04 · Marianne Simonsen, Lars Skipper, Jeffrey A. Smith

This paper uses register-based data to analyze the consequences of a recent major Danish welfare reform for children's academic performance and well-being. In addition to work requirements, the reform brought about consi…

Data augmentation using prosody and false starts to recognize non-native children's speech

2020-08-29 · Hemant Kathania, Mittul Singh, Tamás Grósz, Mikko Kurimo

This paper describes AaltoASR's speech recognition system for the INTERSPEECH 2020 shared task on Automatic Speech Recognition (ASR) for non-native children's speech. The task is to recognize non-native speech from child…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Data AugmentationLanguage Modeling+3

Mini Minds: Exploring Bebeshka and Zlata Baby Models

2023-11-06 · Irina Proskurina, Guillaume Metzler, Julien Velcin

In this paper, we describe the University of Lyon 2 submission to the Strict-Small track of the BabyLM competition. The shared task is created with an emphasis on small-scale language modelling from scratch on limited-si…

DecoderLanguage AcquisitionLanguage Modelling

Econometric model of children participation in family dairy farming in the center of dairy farming, West Java Province, Indonesia

2021-02-05 · Achmad Firman, Ratna Ayu Saptati

The involvement of children in the family dairy farming is pivotal point to reduce the cost of production input, especially in smallholder dairy farming. The purposes of the study are to analysis the factors that influen…