paper-with-me

홈 › Papers

Dynamically Composing Domain-Data Selection with Clean-Data Selection by "Co-Curricular Learning" for Neural Machine Translation

2019-06-03 · Wei Wang, Isaac Caswell, Ciprian Chelba

Noise and domain are important aspects of data quality for neural machine translation. Existing research focus separately on domain-data selection, clean-data selection, or their static combination, leaving the dynamic interaction across them not explicitly examined. This paper introduces a "co-curricular learning" method to compose dynamic domain-data selection with dynamic clean-data selection, for transfer learning across both capabilities. We apply an EM-style optimization procedure to further refine the "co-curriculum". Experiment results and analysis with two domains demonstrate the effectiveness of the method and the properties of data scheduled by the co-curriculum.

📄 PDF Abstract BibTeX arXiv:1906.01130

Code (0)

등록된 구현이 없습니다.

Tasks

Machine TranslationTransfer LearningTranslation

Similar Papers 제목 키워드 기반

Dynamically Composing Domain-Data Selection with Clean-Data Selection by ``Co-Curricular Learning'' for Neural Machine Translation

2019-07-01 · ACL 2019 7 · Wei Wang, Isaac Caswell, Ciprian Chelba

Noise and domain are important aspects of data quality for neural machine translation. Existing research focus separately on domain-data selection, clean-data selection, or their static combination, leaving the dynamic i…

Machine TranslationTransfer LearningTranslation

ProMix: Combating Label Noise via Maximizing Clean Sample Utility

2022-07-21 · Ruixuan Xiao, Yiwen Dong, Haobo Wang, Lei Feng 외

Learning with Noisy Labels (LNL) has become an appealing topic, as imperfectly annotated data are relatively cheaper to obtain. Recent state-of-the-art approaches employ specific selection mechanisms to separate clean an…

Learning with noisy labels

SpectR: Dynamically Composing LM Experts with Spectral Routing

2025-04-04 · William Fleshman, Benjamin Van Durme

Training large, general-purpose language models poses significant challenges. The growing availability of specialized expert models, fine-tuned from pretrained models for specific tasks or domains, offers a promising alt…

Steer2Adapt: Dynamically Composing Steering Vectors Elicits Efficient Adaptation of LLMs

2026-02-07 · Pengrui Han, Xueqiang Xu, Keyang Xuan, Peiyang Song 외 arxiv

Activation steering has emerged as a promising approach for efficiently adapting large language models (LLMs) to downstream behaviors. However, most existing steering methods rely on a single static direction per task or…

HalluClean: A Unified Framework to Combat Hallucinations in LLMs

2025-11-12 · Yaxin Zhao, Yu Zhang arxiv

Large language models (LLMs) have achieved impressive performance across a wide range of natural language processing tasks, yet they often produce hallucinated content that undermines factual reliability. To address this…

Zero-shot GeneralizationQuestion Answering