paper-with-me

Papers

Dynamic Dual Trainable Bounds for Ultra-low Precision Super-Resolution Networks

2022-03-08 · Yunshan Zhong, Mingbao Lin, Xunchao Li, Ke Li, Yunhang Shen, Fei Chao, Yongjian Wu, Rongrong Ji

Light-weight super-resolution (SR) models have received considerable attention for their serviceability in mobile devices. Many efforts employ network quantization to compress SR models. However, these methods suffer from severe performance degradation when quantizing the SR models to ultra-low precision (e.g., 2-bit and 3-bit) with the low-cost layer-wise quantizer. In this paper, we identify that the performance drop comes from the contradiction between the layer-wise symmetric quantizer and the highly asymmetric activation distribution in SR models. This discrepancy leads to either a waste on the quantization levels or detail loss in reconstructed images. Therefore, we propose a novel activation quantizer, referred to as Dynamic Dual Trainable Bounds (DDTB), to accommodate the asymmetry of the activations. Specifically, DDTB innovates in: 1) A layer-wise quantizer with trainable upper and lower bounds to tackle the highly asymmetric activations. 2) A dynamic gate controller to adaptively adjust the upper and lower bounds at runtime to overcome the drastically varying activation ranges over different samples.To reduce the extra overhead, the dynamic gate controller is quantized to 2-bit and applied to only part of the SR networks according to the introduced dynamic intensity. Extensive experiments demonstrate that our DDTB exhibits significant performance improvements in ultra-low precision. For example, our DDTB achieves a 0.70dB PSNR increase on Urban100 benchmark when quantizing EDSR to 2-bit and scaling up output images to x4. Code is at \url{https://github.com/zysxmu/DDTB}.

📄 PDF Abstract BibTeX arXiv:2203.03844

Code (1)

zysxmu/ddtb 공식 구현 pytorch

Tasks

QuantizationSuper-Resolution

Similar Papers 제목 키워드 기반

UltraLIF: Fully Differentiable Spiking Neural Networks via Ultradiscretization and Max-Plus Algebra

2026-02-10 · Jose Marie Antonio Miñoza arxiv

Spiking Neural Networks (SNNs) offer energy-efficient, biologically plausible computation but suffer from non-differentiable spike generation, necessitating reliance on heuristic surrogate gradients. This paper introduce…

Precision Gating: Improving Neural Network Efficiency with Dynamic Dual-Precision Activations

2020-02-17 · ICLR 2020 1 · Yichi Zhang, Ritchie Zhao, Weizhe Hua, Nayun Xu 외

We propose precision gating (PG), an end-to-end trainable dynamic dual-precision quantization technique for deep neural networks. PG computes most features in a low precision and only a small proportion of important feat…

Quantization

DAISS: Phase-Aware Imitation Learning for Dual-Arm Robotic Ultrasound-Guided Interventions

2026-03-08 · Feng Li, Pei Liu, Shiting Wang, Ning Wang 외 arxiv

Imitation learning has shown strong potential for automating complex robotic manipulation. In medical robotics, ultrasound-guided needle insertion demands precise bimanual coordination, as clinicians must simultaneously …

Thermodynamic bounds on ultrasensitivity in covalent switching

2022-11-18 · Jeremy A. Owen, Pranay Talla, John W. Biddle, Jeremy Gunawardena

Switch-like motifs are among the basic building blocks of biochemical networks. A common motif that can serve as an ultrasensitive switch consists of two enzymes acting antagonistically on a substrate, one making and the…

High performance ultra-low-precision convolutions on mobile devices

2017-12-06 · Andrew Tulloch, Yangqing Jia

Many applications of mobile deep learning, especially real-time computer vision workloads, are constrained by computation power. This is particularly true for workloads running on older consumer phones, where a typical d…

CPUDeep LearningVocal Bursts Intensity Prediction