paper-with-me

홈 › Papers

Quantizing Diffusion Models from a Sampling-Aware Perspective

2025-05-04 · Qian Zeng, Jie Song, Yuanyu Wan, Huiqiong Wang, Mingli Song

Diffusion models have recently emerged as the dominant approach in visual generation tasks. However, the lengthy denoising chains and the computationally intensive noise estimation networks hinder their applicability in low-latency and resource-limited environments. Previous research has endeavored to address these limitations in a decoupled manner, utilizing either advanced samplers or efficient model quantization techniques. In this study, we uncover that quantization-induced noise disrupts directional estimation at each sampling step, further distorting the precise directional estimations of higher-order samplers when solving the sampling equations through discretized numerical methods, thereby altering the optimal sampling trajectory. To attain dual acceleration with high fidelity, we propose a sampling-aware quantization strategy, wherein a Mixed-Order Trajectory Alignment technique is devised to impose a more stringent constraint on the error bounds at each sampling step, facilitating a more linear probability flow. Extensive experiments on sparse-step fast sampling across multiple datasets demonstrate that our approach preserves the rapid convergence characteristics of high-speed samplers while maintaining superior generation quality. Code will be made publicly available soon.

📄 PDF Abstract BibTeX arXiv:2505.02242

Code (0)

등록된 구현이 없습니다.

Tasks

DenoisingNoise EstimationQuantization

Similar Papers 제목 키워드 기반

DGQ: Distribution-Aware Group Quantization for Text-to-Image Diffusion Models

2025-01-08 · Hyogon Ryu, Nahyeon Park, Hyunjung Shim

Despite the widespread use of text-to-image diffusion models across various tasks, their computational and memory demands limit practical applications. To mitigate this issue, quantization of diffusion models has been ex…

Quantization

An Analysis on Quantizing Diffusion Transformers

2024-06-16 · Yuewei Yang, Jialiang Wang, Xiaoliang Dai, Peizhao Zhang 외

Diffusion Models (DMs) utilize an iterative denoising process to transform random noise into synthetic data. Initally proposed with a UNet structure, DMs excel at producing images that are virtually indistinguishable wit…

Conditional Image GenerationDenoisingImage GenerationQuantization

AccuQuant: Simulating Multiple Denoising Steps for Quantizing Diffusion Models

2025-10-23 · Seunghoon Lee, Jeongwoo Choi, Byunggwan Son, Jaehyeon Moon 외 arxiv

We present in this paper a novel post-training quantization (PTQ) method, dubbed AccuQuant, for diffusion models. We show analytically and empirically that quantization errors for diffusion models are accumulated over de…

Q-ARVD: Quantizing Autoregressive Video Diffusion Models

2026-05-20 · Siao Tang, Xinyin Ma, Gongfan Fang, Xingyi Yang 외 arxiv

Autoregressive video diffusion models (ARVDs) have emerged as a promising architecture for streaming video generation, paving the way for real-time interactive video generation and world modeling. Despite their potential…

Video Generation

Boosting Diffusion Model for Spectrogram Up-sampling in Text-to-speech: An Empirical Study

2024-06-07 · Chong Zhang, Yanqing Liu, Yang Zheng, Sheng Zhao

Scaling text-to-speech (TTS) with autoregressive language model (LM) to large-scale datasets by quantizing waveform into discrete speech tokens is making great progress to capture the diversity and expressiveness in huma…

DiversityLanguage ModelingLanguage Modellingtext-to-speech+1