paper-with-me

홈 › Papers

Selecting Samples on Graphs: A Unified Dataset Pruning Framework for Lossless Training Acceleration

2026-06-11 · Dongyue Wu, Zilin Guo, Xiaoyu Li, Jiajia Liu, Jingdong Chen, Nong Sang, Changxin Gao arxiv

The rapid growth of modern training datasets has significantly increased computational cost, motivating dataset pruning~(DP) methods which retain only a subset of informative samples to reduce training cost. Existing pruning criteria typically rely on either intrinsic signals that assess samples independently or extrinsic signals that promote diversity via pairwise relations. While effective in their own specific regimes, each captures only one aspect of sample utility and lacks robustness across different pruning ratios or data distribution. In this work, we present a unified graph-based DP framework. By modeling the dataset as a weighted graph, where node weights encode intrinsic value and edge weights encode extrinsic value, DP can be cast as a Maximum Weight Clique Problem (MWCP). Although MWCP is NP-hard, its structure admits a principled greedy solution based on sample-wise marginal gains. Under a few mild conditions, we further prove that this unified objective enjoys a formal approximation guarantee, which applies to a broad family of importance metrics and provides practical design guidelines. Extensive experiments show that our method outperforms existing DP methods while substantially reducing training cost, reducing training time by over 40\% without sacrificing accuracy on ImageNet-1k with ResNet-50.

📄 PDF Abstract BibTeX arXiv:2606.12913

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Model Compression using Progressive Channel Pruning

2025-07-07 · Jinyang Guo, Weichen Zhang, Wanli Ouyang, Dong Xu arxiv

In this work, we propose a simple but effective channel pruning framework called Progressive Channel Pruning (PCP) to accelerate Convolutional Neural Networks (CNNs). In contrast to the existing channel pruning methods t…

Model CompressionTransfer Learning

GDeR: Safeguarding Efficiency, Balancing, and Robustness via Prototypical Graph Pruning

2024-10-17 · Guibin Zhang, Haonan Dong, Yuchen Zhang, ZHIXUN LI 외

Training high-quality deep models necessitates vast amounts of data, resulting in overwhelming computational and memory demands. Recently, data pruning, distillation, and coreset selection have been developed to streamli…

Graph Embedding

RCAP: Robust, Class-Aware, Probabilistic Dynamic Dataset Pruning

2026-06-10 · Atif Hassan, Swanand Khare, Jiaul H. Paik arxiv

Dynamic data pruning techniques aim to reduce computational cost while minimizing information loss by periodically selecting representative subsets of input data during model training. However, existing methods often str…

Transfer Learning

Learning the Network of Graphs for Graph Neural Networks

2022-10-08 · Yixiang Shan, Jielong Yang, Xing Liu, Yixing Gao 외

Graph neural networks (GNNs) have achieved great success in many scenarios with graph-structured data. However, in many real applications, there are three issues when applying GNNs: graphs are unknown, nodes have noisy f…

Graph Neural NetworkRelationRelation Network

JointLK: Joint Reasoning with Language Models and Knowledge Graphs for Commonsense Question Answering

2022-01-16 · ACL ARR January 2022 1 · Anonymous

Existing KG-augmented models for question answering primarily focus on designing elaborate Graph Neural Networks (GNNs) to model knowledge graphs (KGs). However, they ignore (i) the effectively fusing and reasoning over …

Knowledge GraphsQuestion Answering