paper-with-me

홈 › Papers

Data-Efficient Training by Evolved Sampling

2025-09-27 · Ziheng Cheng, Zhong Li, Jiang Bian arxiv

Data selection is designed to accelerate learning with preserved performance. To achieve this, a fundamental thought is to identify informative data samples with significant contributions to the training. In this work, we propose \textbf{Evolved Sampling} (\textbf{ES}), a simple yet effective framework for \emph{dynamic} sampling along the training process. This method conducts \em batch \em level data selection based on the dynamics of losses and augmented \emph{loss differences}, which enables flexible \emph{frequency tuning}, and hence significantly reduces the back propagation time with maintained model performance. Due to its conciseness, ES is also readily extensible to incorporate \em set \em level data selection (to form ES with pruning, \textbf{ESWP}) for further accelerations. As a plug-and-play framework, ES(WP) consistently achieves lossless training accelerations across various pre-training and post-training tasks, saving up to nearly 45\% wall-clock time. Our results motivate further investigations on the data efficiency aspect of modern large-scale machine learning.

📄 PDF Abstract BibTeX arXiv:2509.23461

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Toward Effective Tool-Integrated Reasoning via Self-Evolved Preference Learning

2025-09-27 · Yifei Chen, Guanting Dong, Zhicheng Dou arxiv

Tool-Integrated Reasoning (TIR) enables large language models (LLMs) to improve their internal reasoning ability by integrating external tools. However, models employing TIR often display suboptimal behaviors, such as in…

Human peripheral blur is optimal for object recognition

2018-07-23 · R. T. Pramod, Harish Katti, S. P. Arun

Our vision is sharpest at the center of our gaze and becomes progressively blurry into the periphery. It is widely believed that this high foveal resolution evolved at the expense of peripheral acuity. But what if this s…

ObjectObject Recognition

Evolutionarily-Curated Curriculum Learning for Deep Reinforcement Learning Agents

2019-01-16 · Michael Cerny Green, Benjamin Sergent, Pushyami Shandilya, Vibhor Kumar

In this paper we propose a new training loop for deep reinforcement learning agents with an evolutionary generator. Evolutionary procedural content generation has been used in the creation of maps and levels for games be…

Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Personality Requires Struggle: Three Regimes of the Baldwin Effect in Neuroevolved Chess Agents

2026-04-04 · Diego Armando Resendez Prado arxiv

Can lifetime learning expand behavioral diversity over evolutionary time, rather than collapsing it? Prior theory predicts that plasticity reduces variance by buffering organisms against environmental noise. We test this…

BenchEvolver: Frontier Task Synthesis via Solution-Centric Evolution

2026-05-31 · Yangzhen Wu, Aaron J. Li, Wenjie Ma, Li Cao 외 arxiv

The rapid progress of frontier large language models has led to widespread benchmark saturation, limiting the ability of existing datasets to differentiate model capabilities or provide useful training signal. For instan…