paper-with-me

홈 › Papers

CHIP: CHannel Independence-based Pruning for Compact Neural Networks

2021-10-26 · NeurIPS 2021 12 · Yang Sui, Miao Yin, Yi Xie, Huy Phan, Saman Zonouz, Bo Yuan

Filter pruning has been widely used for neural network compression because of its enabled practical acceleration. To date, most of the existing filter pruning works explore the importance of filters via using intra-channel information. In this paper, starting from an inter-channel perspective, we propose to perform efficient filter pruning using Channel Independence, a metric that measures the correlations among different feature maps. The less independent feature map is interpreted as containing less useful information$/$knowledge, and hence its corresponding filter can be pruned without affecting model capacity. We systematically investigate the quantification metric, measuring scheme and sensitiveness$/$reliability of channel independence in the context of filter pruning. Our evaluation results for different models on various datasets show the superior performance of our approach. Notably, on CIFAR-10 dataset our solution can bring $0.90\%$ and $0.94\%$ accuracy increase over baseline ResNet-56 and ResNet-110 models, respectively, and meanwhile the model size and FLOPs are reduced by $42.8\%$ and $47.4\%$ (for ResNet-56) and $48.3\%$ and $52.1\%$ (for ResNet-110), respectively. On ImageNet dataset, our approach can achieve $40.8\%$ and $44.8\%$ storage and computation reductions, respectively, with $0.15\%$ accuracy increase over the baseline ResNet-50 model. The code is available at https://github.com/Eclipsess/CHIP_NeurIPS2021.

📄 PDF Abstract BibTeX arXiv:2110.13981

Code (1)

eclipsess/chip_neurips2021 공식 구현 pytorch

Tasks

Neural Network Compression

Methods 이 논문이 사용한 방법론

Pruning 설명 없음

Similar Papers 제목 키워드 기반

COMPACT: Common-token Optimized Model Pruning Across Channels and Tokens

2025-09-08 · Eugene Kwek, Wenpeng Yin arxiv

Making large language models (LLMs) more efficient in memory, latency, and serving cost is crucial for edge deployment, interactive applications, and sustainable inference at scale. Pruning is a promising technique, but …

Dynamic Structure Pruning for Compressing CNNs

2023-03-17 · Jun-Hyung Park, Yeachan Kim, Junho Kim, Joon-Young Choi 외

Structure pruning is an effective method to compress and accelerate neural networks. While filter and channel pruning are preferable to other structure pruning methods in terms of realistic acceleration and hardware comp…

GPU

Conditional Automated Channel Pruning for Deep Neural Networks

2020-09-21 · Yixin Liu, Yong Guo, Zichang Liu, Haohua Liu 외

Model compression aims to reduce the redundancy of deep networks to obtain compact models. Recently, channel pruning has become one of the predominant compression methods to deploy deep models on resource-constrained dev…

Model Compression

Lean Unet: A Compact Model for Image Segmentation

2025-12-03 · Ture Hassler, Ida Åkerholm, Marcus Nordström, Gabriele Balletti 외 arxiv

Unet and its variations have been standard in semantic image segmentation, especially for computer assisted radiology. Current Unet architectures iteratively downsample spatial resolution while increasing channel dimensi…

Image Segmentation

OICSR: Out-In-Channel Sparsity Regularization for Compact Deep Neural Networks

2019-05-28 · CVPR 2019 6 · Jiashi Li, Qi Qi, Jingyu Wang, Ce Ge 외

Channel pruning can significantly accelerate and compress deep neural networks. Many channel pruning works utilize structured sparsity regularization to zero out all the weights in some channels and automatically obtain …