paper-with-me

홈 › Papers

Rate-Adaptive Quantization: A Multi-Rate Codebook Adaptation for Vector Quantization-based Generative Models

2024-05-23 · Jiwan Seo, Joonhyuk Kang

Learning discrete representations with vector quantization (VQ) has emerged as a powerful approach in various generative models. However, most VQ-based models rely on a single, fixed-rate codebook, requiring extensive retraining for new bitrates or efficiency requirements. We introduce Rate-Adaptive Quantization (RAQ), a multi-rate codebook adaptation framework for VQ-based generative models. RAQ applies a data-driven approach to generate variable-rate codebooks from a single baseline VQ model, enabling flexible tradeoffs between compression and reconstruction fidelity. Additionally, we provide a simple clustering-based procedure for pre-trained VQ models, offering an alternative when retraining is infeasible. Our experiments show that RAQ performs effectively across multiple rates, often outperforming conventional fixed-rate VQ baselines. By enabling a single system to seamlessly handle diverse bitrate requirements, RAQ extends the adaptability of VQ-based generative models and broadens their applicability to data compression, reconstruction, and generation tasks.

📄 PDF Abstract BibTeX arXiv:2405.14222

Code (1)

JiwanSeo/RAQ-VAE 공식 구현 pytorch

Tasks

Data CompressionImage GenerationImage ReconstructionQuantization

Methods 이 논문이 사용한 방법론

USD Coin Customer Service Number +1-833-534-1729 설명 없음
VQ-VAE VQ-VAE is a type of variational autoencoder that uses vector quantisation to obtain a discrete latent representation. It differs from…

Similar Papers 제목 키워드 기반

AAAC: Activation-Aware Adaptive Codebooks for 4-bit LLM Weight Quantization

2026-05-09 · Beshr IslamBouli, David Jin arxiv

Post-training weight-only quantization to 4 bits is widely used to reduce the memory and compute costs of large language model inference. Existing PTQ methods, such as AWQ and GPTQ, improve how weights are mapped onto a …

Balance of Number of Embedding and their Dimensions in Vector Quantization

2024-07-06 · Hang Chen, Sankepally Sainath Reddy, Ziwei Chen, Dianbo Liu

The dimensionality of the embedding and the number of available embeddings ( also called codebook size) are critical factors influencing the performance of Vector Quantization(VQ), a discretization process used in many m…

Quantization

ESC-MVQ: End-to-End Semantic Communication With Multi-Codebook Vector Quantization

2025-04-16 · Junyong Shin, Yongjeong Oh, Jinsung Park, Joohyuk Park 외

This paper proposes a novel end-to-end digital semantic communication framework based on multi-codebook vector quantization (VQ), referred to as ESC-MVQ. Unlike prior approaches that rely on end-to-end training with a sp…

DecoderQuantizationSemantic Communication

Adaptive Distribution-aware Quantization for Mixed-Precision Neural Networks

2025-10-22 · Shaohang Jia, Zhiyong Huang, Zhi Yu, Mingyang Hou 외 arxiv

Quantization-Aware Training (QAT) is a critical technique for deploying deep neural networks on resource-constrained devices. However, existing methods often face two major challenges: the highly non-uniform distribution…

Not All Image Regions Matter: Masked Vector Quantization for Autoregressive Image Generation

2023-05-23 · CVPR 2023 1 · Mengqi Huang, Zhendong Mao, Quan Wang, Yongdong Zhang

Existing autoregressive models follow the two-stage generation paradigm that first learns a codebook in the latent space for image reconstruction and then completes the image generation autoregressively based on the lear…

AllImage GenerationImage ReconstructionQuantization