paper-with-me

홈 › Papers

Nonparametric Inference under B-bits Quantization

2019-01-24 · Kexuan Li, Ruiqi Liu, Ganggang Xu, Zuofeng Shang

Statistical inference based on lossy or incomplete samples is often needed in research areas such as signal/image processing, medical image storage, remote sensing, signal transmission. In this paper, we propose a nonparametric testing procedure based on samples quantized to $B$ bits through a computationally efficient algorithm. Under mild technical conditions, we establish the asymptotic properties of the proposed test statistic and investigate how the testing power changes as $B$ increases. In particular, we show that if $B$ exceeds a certain threshold, the proposed nonparametric testing procedure achieves the classical minimax rate of testing (Shang and Cheng, 2015) for spline models. We further extend our theoretical investigations to a nonparametric linearity test and an adaptive nonparametric test, expanding the applicability of the proposed methods. Extensive simulation studies {together with a real-data analysis} are used to demonstrate the validity and effectiveness of the proposed tests.

📄 PDF Abstract BibTeX arXiv:1901.08571

Code (0)

등록된 구현이 없습니다.

Tasks

Quantization

Similar Papers 제목 키워드 기반

Quantized Nonparametric Estimation over Sobolev Ellipsoids

2015-03-25 · Yuancheng Zhu, John Lafferty

We formulate the notion of minimax estimation under storage or communication constraints, and prove an extension to Pinsker's theorem for nonparametric estimation over Sobolev ellipsoids. Placing limits on the number of …

Quantization

On the Expressive Power of Weight Quantization in Large Language Models

2026-06-20 · Shao-Qun Zhang arxiv

In recent years, weight quantization that encodes the learnable parameters of large language models in an $n$-bit format has garnered significant attention due to its potential for model compression and inference acceler…

Model Compression

Extremely Low Bit Transformer Quantization for On-Device Neural Machine Translation

2020-09-16 · Findings of the Association for Computational Linguistics 2020 · Insoo Chung, Byeongwook Kim, Yoonjung Choi, Se Jung Kwon 외

The deployment of widely used Transformer architecture is challenging because of heavy computation load and memory overhead during inference, especially when the target device is limited in computational resources such a…

Machine TranslationNMTQuantizationTranslation

Class-based Quantization for Neural Networks

2022-11-27 · Wenhao Sun, Grace Li Zhang, Huaxi Gu, Bing Li 외

In deep neural networks (DNNs), there are a huge number of weights and multiply-and-accumulate (MAC) operations. Accordingly, it is challenging to apply DNNs on resource-constrained platforms, e.g., mobile phones. Quanti…

Quantization

CCQ: Convolutional Code for Extreme Low-bit Quantization in LLMs

2025-07-09 · Zhaojing Zhou, Xunchao Li, Minghao Li, Handi Zhang 외 arxiv

The rapid scaling of Large Language Models (LLMs) elevates inference costs and compounds substantial deployment barriers. While quantization to 8 or 4 bits mitigates this, sub-3-bit methods face severe accuracy, scalabil…