paper-with-me

Papers

Model compression as constrained optimization, with application to neural nets. Part I: general framework

2017-07-05 · Miguel Á. Carreira-Perpiñán

Compressing neural nets is an active research problem, given the large size of state-of-the-art nets for tasks such as object recognition, and the computational limits imposed by mobile devices. We give a general formulation of model compression as constrained optimization. This includes many types of compression: quantization, low-rank decomposition, pruning, lossless compression and others. Then, we give a general algorithm to optimize this nonconvex problem based on the augmented Lagrangian and alternating optimization. This results in a "learning-compression" algorithm, which alternates a learning step of the uncompressed model, independent of the compression type, with a compression step of the model parameters, independent of the learning task. This simple, efficient algorithm is guaranteed to find the best compressed model for the task in a local sense under standard assumptions. We present separately in several companion papers the development of this general framework into specific algorithms for model compression based on quantization, pruning and other variations, including experimental results on compressing neural nets and other models.

📄 PDF Abstract BibTeX arXiv:1707.01209

Code (1)

UCMerced-ML/LC-model-compression pytorch

Tasks

Model CompressionObject RecognitionQuantization

Methods 이 논문이 사용한 방법론

Pruning 설명 없음

Similar Papers 제목 키워드 기반

Model compression as constrained optimization, with application to neural nets. Part V: combining compressions

2021-07-09 · Miguel Á. Carreira-Perpiñán, Yerlan Idelbayev

Model compression is generally performed by using quantization, low-rank approximation or pruning, for which various algorithms have been researched in recent years. One fundamental question is: what types of compression…

Additive modelsLow-rank compressionModel CompressionNetwork Pruning+1

Model compression as constrained optimization, with application to neural nets. Part II: quantization

2017-07-13 · Miguel Á. Carreira-Perpiñán, Yerlan Idelbayev

We consider the problem of deep neural net compression by quantization: given a large, reference net, we want to quantize its real-valued weights using a codebook with $K$ entries so that the training loss of the quantiz…

BinarizationModel CompressionQuantization

EAST: Encoding-Aware Sparse Training for Deep Memory Compression of ConvNets

2019-12-20 · Matteo Grimaldi, Valentino Peluso, Andrea Calimera

The implementation of Deep Convolutional Neural Networks (ConvNets) on tiny end-nodes with limited non-volatile memory space calls for smart compression strategies capable of shrinking the footprint yet preserving predic…

Quantization

Wide Compression: Tensor Ring Nets

2018-02-25 · CVPR 2018 6 · Wenqi Wang, Yifan Sun, Brian Eriksson, Wenlin Wang 외

Deep neural networks have demonstrated state-of-the-art performance in a variety of real-world applications. In order to obtain performance gains, these networks have grown larger and deeper, containing millions or even …

image-classificationImage Classification

Constrained Deep Learning using Conditional Gradient and Applications in Computer Vision

2018-03-17 · Sathya N. Ravi, Tuan Dinh, Vishnu Sai Rao Lokhande, Vikas Singh

A number of results have recently demonstrated the benefits of incorporating various constraints when training deep architectures in vision and machine learning. The advantages range from guarantees for statistical gener…

Image Inpainting