paper-with-me

홈 › Papers

Detecting Dead Weights and Units in Neural Networks

2018-06-15 · Utku Evci

Deep Neural Networks are highly over-parameterized and the size of the neural networks can be reduced significantly after training without any decrease in performance. One can clearly see this phenomenon in a wide range of architectures trained for various problems. Weight/channel pruning, distillation, quantization, matrix factorization are some of the main methods one can use to remove the redundancy to come up with smaller and faster models. This work starts with a short informative chapter, where we motivate the pruning idea and provide the necessary notation. In the second chapter, we compare various saliency scores in the context of parameter pruning. Using the insights obtained from this comparison and stating the problems it brings we motivate why pruning units instead of the individual parameters might be a better idea. We propose some set of definitions to quantify and analyze units that don't learn and create any useful information. We propose an efficient way for detecting dead units and use it to select which units to prune. We get 5x model size reduction through unit-wise pruning on MNIST.

📄 PDF Abstract BibTeX arXiv:1806.06068

Code (0)

등록된 구현이 없습니다.

Tasks

Quantization

Methods 이 논문이 사용한 방법론

Pruning 설명 없음

Similar Papers 제목 키워드 기반

Synaptic Stripping: How Pruning Can Bring Dead Neurons Back To Life

2023-02-11 · Tim Whitaker, Darrell Whitley

Rectified Linear Units (ReLU) are the default choice for activation functions in deep neural networks. While they demonstrate excellent empirical performance, ReLU activations can fall victim to the dead neuron problem. …

Superseding Model Scaling by Penalizing Dead Units and Points with Separation Constraints

2019-09-25 · Carles Riera, Camilo Rey-Torres, Eloi Puertas, Oriol Pujol

In this article, we study a proposal that enables to train extremely thin (4 or 8 neurons per layer) and relatively deep (more than 100 layers) feedforward networks without resorting to any architectural modification suc…

ReCU: Reviving the Dead Weights in Binary Neural Networks

2021-03-23 · ICCV 2021 10 · Zihan Xu, Mingbao Lin, Jianzhuang Liu, Jie Chen 외

Binary neural networks (BNNs) have received increasing attention due to their superior reductions of computation and memory. Most existing works focus on either lessening the quantization error by minimizing the gap betw…

BinarizationQuantization

Mean Replacement Pruning

2019-05-01 · ICLR 2019 5 · Utku Evci, Nicolas Le Roux, Pablo Castro, Leon Bottou

Pruning units in a deep network can help speed up inference and training as well as reduce the size of the model. We show that bias propagation is a pruning technique which consistently outperforms the common approach of…

Good Weights: Proactive, Adaptive Dead Reckoning Fusion for Continuous and Robust Visual SLAM

2025-09-26 · Yanwei Du, Jing-Chen Peng, Patricio A. Vela arxiv

Given that Visual SLAM relies on appearance cues for localization and scene understanding, texture-less or visually degraded environments (e.g., plain walls or low lighting) lead to poor pose estimation and track loss. H…

Scene UnderstandingVisual TrackingPose Estimation