paper-with-me

Papers

Multi-Step Consistency Models: Fast Generation with Theoretical Guarantees

2025-05-02 · Nishant Jain, Xunpeng Huang, Yian Ma, Tong Zhang

Consistency models have recently emerged as a compelling alternative to traditional SDE-based diffusion models. They offer a significant acceleration in generation by producing high-quality samples in very few steps. Despite their empirical success, a proper theoretic justification for their speed-up is still lacking. In this work, we address the gap by providing a theoretical analysis of consistency models capable of mapping inputs at a given time to arbitrary points along the reverse trajectory. We show that one can achieve a KL divergence of order $ O(\varepsilon^2) $ using only $ O\left(\log\left(\frac{d}{\varepsilon}\right)\right) $ iterations with a constant step size. Additionally, under minimal assumptions on the data distribution (non smooth case) an increasingly common setting in recent diffusion model analyses we show that a similar KL convergence guarantee can be obtained, with the number of steps scaling as $ O\left(d \log\left(\frac{d}{\varepsilon}\right)\right) $. Going further, we also provide a theoretical analysis for estimation of such consistency models, concluding that accurate learning is feasible using small discretization steps, both in smooth and non-smooth settings. Notably, our results for the non-smooth case yield best in class convergence rates compared to existing SDE or ODE based analyses under minimal assumptions.

📄 PDF Abstract BibTeX arXiv:2505.01049

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Consistency Models 설명 없음
Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…
SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…
NON 설명 없음

Similar Papers 제목 키워드 기반

Convergence Of Consistency Model With Multistep Sampling Under General Data Assumptions

2025-05-06 · Yiding Chen, Yiyi Zhang, Owen Oertell, Wen Sun

Diffusion models accomplish remarkable success in data generation tasks across various domains. However, the iterative sampling process is computationally expensive. Consistency models are proposed to learn consistency f…

Flow map matching with stochastic interpolants: A mathematical framework for consistency models

2024-06-11 · Nicholas M. Boffi, Michael S. Albergo, Eric Vanden-Eijnden

Generative models based on dynamical equations such as flows and diffusions offer exceptional sample quality, but require computationally expensive numerical integration during inference. The advent of consistency models…

Numerical Integration

ECTraj: Enhanced Consistency Training for Multi-Agent Trajectory Prediction

2026-05-09 · Alen Mrdovic, Qingze, Liu, Danrui Li 외 arxiv

Diffusion models for multi-agent trajectory prediction are limited by iterative denoising, which causes inference latency that hinders their use in time-critical settings like autonomous driving. Fast-sampling variants u…

Trajectory PredictionAutonomous Driving

FRMD: Fast Robot Motion Diffusion with Consistency-Distilled Movement Primitives for Smooth Action Generation

2025-03-03 · Xirui Shi, Jun Jin

We consider the problem of using diffusion models to generate fast, smooth, and temporally consistent robot motions. Although diffusion models have demonstrated superior performance in robot learning due to their task sc…

Action GenerationDenoisingImage GenerationMotion Generation

Mutual Forcing: Dual-Mode Self-Evolution for Fast Autoregressive Audio-Video Character Generation

2026-04-28 · Yupeng Zhou, Lianghua Huang, Zhifan Wu, Jiabao Wang 외 arxiv

In this work, we propose Mutual Forcing, a framework for fast autoregressive audio-video generation with long-horizon audio-video synchronization. Our approach addresses two key challenges: joint audio-video modeling and…

Video Generation