paper-with-me

홈 › Papers

Leveraging Filter Correlations for Deep Model Compression

2018-11-26 · Pravendra Singh, Vinay Kumar Verma, Piyush Rai, Vinay P. Namboodiri

We present a filter correlation based model compression approach for deep convolutional neural networks. Our approach iteratively identifies pairs of filters with the largest pairwise correlations and drops one of the filters from each such pair. However, instead of discarding one of the filters from each such pair na\"{i}vely, the model is re-optimized to make the filters in these pairs maximally correlated, so that discarding one of the filters from the pair results in minimal information loss. Moreover, after discarding the filters in each round, we further finetune the model to recover from the potential small loss incurred by the compression. We evaluate our proposed approach using a comprehensive set of experiments and ablation studies. Our compression method yields state-of-the-art FLOPs compression rates on various benchmarks, such as LeNet-5, VGG-16, and ResNet-50,56, while still achieving excellent predictive performance for tasks such as object detection on benchmark datasets.

📄 PDF Abstract BibTeX arXiv:1811.10559

Code (0)

등록된 구현이 없습니다.

Tasks

modelModel Compressionobject-detectionObject Detection

Similar Papers 제목 키워드 기반

Edge-Fog Computing-Enabled EEG Data Compression via Asymmetrical Variational Discrete Cosine Transform Network

2025-03-13 · Xin Zhu, Hongyi Pan, Ahmet Enis Cetin

The large volume of electroencephalograph (EEG) data produced by brain-computer interface (BCI) systems presents challenges for rapid transmission over bandwidth-limited channels in Internet of Things (IoT) networks. To …

Brain Computer InterfaceData CompressionEEGVariational Inference

Q-Filters: Leveraging QK Geometry for Efficient KV Cache Compression

2025-03-04 · Nathan Godey, Alessio Devoto, Yu Zhao, Simone Scardapane 외

Autoregressive language models rely on a Key-Value (KV) Cache, which avoids re-computing past hidden states during generation, making it faster. As model sizes and context lengths grow, the KV Cache becomes a significant…

Text Generation

ECVC: Exploiting Non-Local Correlations in Multiple Frames for Contextual Video Compression

2024-10-13 · CVPR 2025 1 · Wei Jiang, Junru Li, Kai Zhang, Li Zhang

In Learned Video Compression (LVC), improving inter prediction, such as enhancing temporal context mining and mitigating accumulated errors, is crucial for boosting rate-distortion performance. Existing LVCs mainly focus…

Video Compression

Compressing gradients by exploiting temporal correlation in momentum-SGD

2021-08-17 · Tharindu B. Adikari, Stark C. Draper

An increasing bottleneck in decentralized optimization is communication. Bigger models and growing datasets mean that decentralization of computation is important and that the amount of information exchanged is quickly g…

DynaFilter: Cloud-driven Dynamic Filtering for Satellite Edge Intelligence

2026-07-11 · Ziyang Zhang, Jie Liu, Luca Mottola arxiv

Modern satellite edge systems, including those performing remote sensing tasks such object detection and tracking, are characterized by severely limited bandwidth and intermittent connections, making continuous data tran…

Object Detection