paper-with-me

홈 › Papers

Probabilistic fine-tuning of pruning masks and PAC-Bayes self-bounded learning

2021-10-22 · Soufiane Hayou, Bobby He, Gintare Karolina Dziugaite

We study an approach to learning pruning masks by optimizing the expected loss of stochastic pruning masks, i.e., masks which zero out each weight independently with some weight-specific probability. We analyze the training dynamics of the induced stochastic predictor in the setting of linear regression, and observe a data-adaptive L1 regularization term, in contrast to the dataadaptive L2 regularization term known to underlie dropout in linear regression. We also observe a preference to prune weights that are less well-aligned with the data labels. We evaluate probabilistic fine-tuning for optimizing stochastic pruning masks for neural networks, starting from masks produced by several baselines. In each case, we see improvements in test error over baselines, even after we threshold fine-tuned stochastic pruning masks. Finally, since a stochastic pruning mask induces a stochastic neural network, we consider training the weights and/or pruning probabilities simultaneously to minimize a PAC-Bayes bound on generalization error. Using data-dependent priors, we obtain a selfbounded learning algorithm with strong performance and numerically tight bounds. In the linear model, we show that a PAC-Bayes generalization error bound is controlled by the magnitude of the change in feature alignment between the 'prior' and 'posterior' data.

📄 PDF Abstract BibTeX arXiv:2110.11804

Code (0)

등록된 구현이 없습니다.

Tasks

L2 Regularizationregression

Methods 이 논문이 사용한 방법론

Test 설명 없음
Pruning 설명 없음
Dropout Dropout is a regularization technique for neural networks that drops a unit (along with connections) at training time with a specified probability $p$ (a common value is…
L1 Regularization $L_{1}$ Regularization is a regularization technique applied to the weights of a neural network. We minimize a loss function compromising both the primary loss function and a…

Similar Papers 제목 키워드 기반

QPruner: Probabilistic Decision Quantization for Structured Pruning in Large Language Models

2024-12-16 · Changhai Zhou, Yuhua Zhou, Shijie Han, Qian Qiao 외

The rise of large language models (LLMs) has significantly advanced various natural language processing (NLP) tasks. However, the resource demands of these models pose substantial challenges. Structured pruning is an eff…

Bayesian OptimizationQuantization

Bypass Back-propagation: Optimization-based Structural Pruning for Large Language Models via Policy Gradient

2024-06-15 · Yuan Gao, Zujing Liu, Weizhong Zhang, Bo Du 외

In contrast to moderate-size neural network pruning, structural weight pruning on the Large-Language Models (LLMs) imposes a novel challenge on the efficiency of the pruning algorithms, due to the heavy computation/memor…

GPUNetwork Pruning

Stochastic Subnetwork Annealing: A Regularization Technique for Fine Tuning Pruned Subnetworks

2024-01-16 · Tim Whitaker, Darrell Whitley

Pruning methods have recently grown in popularity as an effective way to reduce the size and computational complexity of deep neural networks. Large numbers of parameters can be removed from trained models with little di…

Weight Reparametrization for Budget-Aware Network Pruning

2021-07-08 · Robin Dupont, Hichem Sahbi, Guillaume Michel

Pruning seeks to design lightweight architectures by removing redundant weights in overparameterized networks. Most of the existing techniques first remove structured sub-networks (filters, channels,...) and then fine-tu…

Network Pruning

Pruning-aware Sparse Regularization for Network Pruning

2022-01-18 · Nanfei Jiang, Xu Zhao, Chaoyang Zhao, Yongqi An 외

Structural neural network pruning aims to remove the redundant channels in the deep convolutional neural networks (CNNs) by pruning the filters of less importance to the final output accuracy. To reduce the degradation o…

Network Pruning