paper-with-me

홈 › Papers

Dynamic Jointly Batch Selection for Data Efficient Machine Translation Fine-Tuning

2025-11-06 · Mohammad Amin Ghanizadeh, Mohammad Javad Dousti arxiv

Data quality and its effective selection are fundamental to improving the performance of machine translation models, serving as cornerstones for achieving robust and reliable translation systems. This paper presents a data selection methodology specifically designed for fine-tuning machine translation systems, which leverages the synergy between a learner model and a pre-trained reference model to enhance overall training effectiveness. By defining a learnability score, our approach systematically evaluates the utility of data points for training, ensuring that only the most relevant and impactful examples contribute to the fine-tuning process. Furthermore, our method employs a batch selection strategy which considers interdependencies among data points, optimizing the efficiency of the training process while maintaining a focus on data relevance. Experiments on English to Persian and several other language pairs using an mBART model fine-tuned on the CCMatrix dataset demonstrate that our method can achieve up to a fivefold improvement in data efficiency compared to an iid baseline. Experimental results indicate that our approach improves computational efficiency by 24 when utilizing cached embeddings, as it requires fewer training data points. Additionally, it enhances generalization, resulting in superior translation performance compared to random selection method.

📄 PDF Abstract BibTeX arXiv:2511.04406

Code (0)

등록된 구현이 없습니다.

Tasks

Computational EfficiencyMachine Translation

Similar Papers 제목 키워드 기반

scikit-dyn2sel -- A Dynamic Selection Framework for Data Streams

2020-08-17 · Lucca Portes Cavalheiro, Jean Paul Barddal, Alceu de Souza Britto Jr, Laurent Heutte

Mining data streams is a challenge per se. It must be ready to deal with an enormous amount of data and with problems not present in batch machine learning, such as concept drift. Therefore, applying a batch-designed tec…

Optimizing Data Curation through Spectral Analysis and Joint Batch Selection (SALN)

2024-12-22 · Mohammadreza Sharifi

In modern deep learning models, long training times and large datasets present significant challenges to both efficiency and scalability. Effective data curation and sample selection are crucial for optimizing the traini…

D2ACE: Multi-Label Batch Selection Guided by Dual Dynamics and Adaptive Correlation Enhancement

2026-05-10 · Bin Liu, Haoyu Peng, Zhijia Wei, Jiajing Zhang 외 arxiv

Batch selection is crucial for improving both training efficiency and predictive performance in deep multi-label classification (MLC). Existing batch selection methods typically rely on a single metric to assess instance…

Multi-Label Classification

FairBatch: Batch Selection for Model Fairness

2020-12-03 · ICLR 2021 1 · Yuji Roh, Kangwook Lee, Steven Euijong Whang, Changho Suh

Training a fair machine learning model is essential to prevent demographic disparity. Existing techniques for improving model fairness require broad changes in either data preprocessing or model training, rendering thems…

BIG-bench Machine LearningBilevel OptimizationFairnessmodel

Accelerating Minibatch Stochastic Gradient Descent using Typicality Sampling

2019-03-11 · Xinyu Peng, Li Li, Fei-Yue Wang

Machine learning, especially deep neural networks, has been rapidly developed in fields including computer vision, speech recognition and reinforcement learning. Although Mini-batch SGD is one of the most popular stochas…

Reinforcement LearningReinforcement Learning (RL)speech-recognitionSpeech Recognition+1