paper-with-me

Papers

Approximation speed of quantized vs. unquantized ReLU neural networks and beyond

2022-05-24 · Antoine Gonon, Nicolas Brisebarre, Rémi Gribonval, Elisa Riccietti

We deal with two complementary questions about approximation properties of ReLU networks. First, we study how the uniform quantization of ReLU networks with real-valued weights impacts their approximation properties. We establish an upper-bound on the minimal number of bits per coordinate needed for uniformly quantized ReLU networks to keep the same polynomial asymptotic approximation speeds as unquantized ones. We also characterize the error of nearest-neighbour uniform quantization of ReLU networks. This is achieved using a new lower-bound on the Lipschitz constant of the map that associates the parameters of ReLU networks to their realization, and an upper-bound generalizing classical results. Second, we investigate when ReLU networks can be expected, or not, to have better approximation properties than other classical approximation families. Indeed, several approximation families share the following common limitation: their polynomial asymptotic approximation speed of any set is bounded from above by the encoding speed of this set. We introduce a new abstract property of approximation families, called infinite-encodability, which implies this upper-bound. Many classical approximation families, defined with dictionaries or ReLU networks, are shown to be infinite-encodable. This unifies and generalizes several situations where this upper-bound is known.

📄 PDF Abstract BibTeX arXiv:2205.11874

Code (0)

등록된 구현이 없습니다.

Tasks

Quantization

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

On the Universal Approximability and Complexity Bounds of Quantized ReLU Neural Networks

2018-02-10 · ICLR 2019 5 · Yukun Ding, Jinglan Liu, JinJun Xiong, Yiyu Shi

Compression is a key step to deploy large neural networks on resource-constrained platforms. As a popular compression technique, quantization constrains the number of distinct weight values and thus reducing the number o…

Quantization

Understanding the Impact of Post-Training Quantization on Large Language Models

2023-09-11 · Somnath Roy

Large language models (LLMs) are rapidly increasing in size, with the number of parameters becoming a key factor in the success of many commercial models, such as ChatGPT, Claude, and Bard. Even the recently released pub…

Quantization

On the Effect of Quantization on Dynamic Mode Decomposition

2024-04-02 · Dipankar Maity, Debdipta Goswami, Sriram Narayanan

Dynamic Mode Decomposition (DMD) is a widely used data-driven algorithm for estimating the Koopman Operator.This paper investigates how the estimation process is affected when the data is quantized. Specifically, we exam…

Quantization

Differentially Quantized Gradient Methods

2020-02-06 · Chung-Yi Lin, Victoria Kostina, Babak Hassibi

Consider the following distributed optimization scenario. A worker has access to training data that it uses to compute the gradients while a server decides when to stop iterative computation based on its target accuracy …

Distributed OptimizationQuantization

Koopman Meets Limited Bandwidth: Effect of Quantization on Data-Driven Linear Prediction and Control of Nonlinear Systems

2025-01-13 · Shahab Ataei, Dipankar Maity, Debdipta Goswami

Koopman-based lifted linear identification have been widely used for data-driven prediction and model predictive control (MPC) of nonlinear systems. It has found applications in flow-control, soft robotics, and unmanned …

Model Predictive ControlQuantization