paper-with-me

Papers

Texture Vector-Quantization and Reconstruction Aware Prediction for Generative Super-Resolution

2025-09-28 · Qifan Li, Jiale Zou, Jinhua Zhang, Wei Long, Xingyu Zhou, Shuhang Gu arxiv

Vector-quantized based models have recently demonstrated strong potential for visual prior modeling. However, existing VQ-based methods simply encode visual features with nearest codebook items and train index predictor with code-level supervision. Due to the richness of visual signal, VQ encoding often leads to large quantization error. Furthermore, training predictor with code-level supervision can not take the final reconstruction errors into consideration, result in sub-optimal prior modeling accuracy. In this paper we address the above two issues and propose a Texture Vector-Quantization and a Reconstruction Aware Prediction strategy. The texture vector-quantization strategy leverages the task character of super-resolution and only introduce codebook to model the prior of missing textures. While the reconstruction aware prediction strategy makes use of the straight-through estimator to directly train index predictor with image-level supervision. Our proposed generative SR model (TVQ&RAP) is able to deliver photo-realistic SR results with small computational cost.

📄 PDF Abstract BibTeX arXiv:2509.23774

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

VAR-3D: View-aware Auto-Regressive Model for Text-to-3D Generation via a 3D Tokenizer

2026-02-14 · Zongcheng Han, Dongyan Cao, Haoran Sun, Yu Hong arxiv

Recent advances in auto-regressive transformers have achieved remarkable success in generative modeling. However, text-to-3D generation remains challenging, primarily due to bottlenecks in learning discrete 3D representa…

3D Generation

Two-Dimensional Quantization for Geometry-Aware Audio Coding

2025-12-01 · Tal Shuster, Eliya Nachmani arxiv

Recent neural audio codecs have achieved impressive reconstruction quality, typically relying on quantization methods such as Residual Vector Quantization (RVQ), Vector Quantization (VQ) and Finite Scalar Quantization (F…

Representation Learning

FlowVQTalker: High-Quality Emotional Talking Face Generation through Normalizing Flow and Quantization

2024-03-11 · CVPR 2024 1 · Shuai Tan, Bin Ji, Ye Pan

Generating emotional talking faces is a practical yet challenging endeavor. To create a lifelike avatar, we draw upon two critical insights from a human perspective: 1) The connection between audio and the non-determinis…

Face GenerationQuantizationTalking Face Generation

Generalized Radius and Integrated Codebook Transforms for Differentiable Vector Quantization

2026-02-01 · Haochen You, Heng Zhang, Hongyang He, Yuqi Li 외 arxiv

Vector quantization (VQ) underpins modern generative and representation models by turning continuous latents into discrete tokens. Yet hard nearest-neighbor assignments are non-differentiable and are typically optimized …

Image ReconstructionImage Generation

Channel-Aware Vector Quantization for Robust Semantic Communication on Discrete Channels

2025-10-21 · Zian Meng, Qiang Li, Wenqian Tang, Mingdie Yan 외 arxiv

Deep learning-based semantic communication has largely relied on analog or semi-digital transmission, which limits compatibility with modern digital communication infrastructures. Recent studies have employed vector quan…

Semantic Communication