paper-with-me

홈 › Papers

Optimal Corpus Aware Training for Neural Machine Translation

2025-08-07 · Yi-Hsiu Liao, Cheng Shen, Brenda, Yang arxiv

Corpus Aware Training (CAT) leverages valuable corpus metadata during training by injecting corpus information into each training example, and has been found effective in the literature, commonly known as the "tagging" approach. Models trained with CAT inherently learn the quality, domain and nuance between corpora directly from data, and can easily switch to different inference behavior. To achieve the best evaluation, CAT models pre-define a group of high quality data before training starts which can be error-prone and inefficient. In this work, we propose Optimal Corpus Aware Training (OCAT), which fine-tunes a CAT pre-trained model by freezing most of the model parameters and only tuning small set of corpus-related parameters. We show that OCAT is lightweight, resilient to overfitting, and effective in boosting model accuracy. We use WMT23 English to Chinese and English to German translation tasks as our test ground and show +3.6 and +1.8 chrF improvement, respectively, over vanilla training. Furthermore, our approach is on-par or slightly better than other state-of-the-art fine-tuning techniques while being less sensitive to hyperparameter settings.

📄 PDF Abstract BibTeX arXiv:2508.05364

Code (0)

등록된 구현이 없습니다.

Tasks

Machine Translation

Similar Papers 제목 키워드 기반

Context Based Machine Translation With Recurrent Neural Network For English-Amharic Translation

2019-09-25 · Yeabsira Asefa Ashengo, Rosa Tsegaye Aga, Surafel Lemma Abebe

The current approaches for machine translation usually require large set of parallel corpus in order to achieve fluency like in the case of neural machine translation (NMT), statistical machine translation (SMT) and exam…

Machine TranslationNMTTranslation

TANDO: A Corpus for Document-level Machine Translation

2022-06-01 · LREC 2022 6 · Harritxu Gete, Thierry Etchegoyhen, David Ponce, Gorka Labaka 외

Document-level Neural Machine Translation aims to increase the quality of neural translation models by taking into account contextual information. Properly modelling information beyond the sentence level can result in im…

Document Level Machine TranslationMachine TranslationSentenceTranslation

NeoAMT: Neologism-Aware Agentic Machine Translation with Reinforcement Learning

2026-01-07 · Zhongtao Miao, Kaiyan Zhao, Masaaki Nagata, Yoshimasa Tsuruoka arxiv

Neologism-aware machine translation aims to translate source sentences containing neologisms into target languages. This field remains underexplored compared with general machine translation (MT). In this paper, we propo…

Reinforcement LearningMachine Translation

Approches quantitatives de l'analyse des pr{é}dictions en traduction automatique neuronale (TAN)

2020-12-10 · Maria Zimina-Poirot, Nicolas Ballier, Jean-Baptiste Yunès

As part of a larger project on optimal learning conditions in neural machine translation, we investigate characteristic training phases of translation engines. All our experiments are carried out using OpenNMT-Py: the pr…

Machine TranslationNMTTranslation

Breaking the Corpus Bottleneck for Context-Aware Neural Machine Translation with Cross-Task Pre-training

2021-08-01 · ACL 2021 5 · Linqing Chen, Junhui Li, ZhengXian Gong, Boxing Chen 외

Context-aware neural machine translation (NMT) remains challenging due to the lack of large-scale document-level parallel corpora. To break the corpus bottleneck, in this paper we aim to improve context-aware NMT by taki…

Machine TranslationNMTSentenceTranslation