paper-with-me

Papers

Accelerating Image Generation with Sub-path Linear Approximation Model

2024-04-22 · Chen Xu, Tianhui Song, Weixin Feng, Xubin Li, Tiezheng Ge, Bo Zheng, LiMin Wang

Diffusion models have significantly advanced the state of the art in image, audio, and video generation tasks. However, their applications in practical scenarios are hindered by slow inference speed. Drawing inspiration from the approximation strategies utilized in consistency models, we propose the Sub-path Linear Approximation Model (SLAM), which accelerates diffusion models while maintaining high-quality image generation. SLAM treats the PF-ODE trajectory as a series of PF-ODE sub-paths divided by sampled points, and harnesses sub-path linear (SL) ODEs to form a progressive and continuous error estimation along each individual PF-ODE sub-path. The optimization on such SL-ODEs allows SLAM to construct denoising mappings with smaller cumulative approximated errors. An efficient distillation method is also developed to facilitate the incorporation of more advanced diffusion models, such as latent diffusion models. Our extensive experimental results demonstrate that SLAM achieves an efficient training regimen, requiring only 6 A100 GPU days to produce a high-quality generative model capable of 2 to 4-step generation with high performance. Comprehensive evaluations on LAION, MS COCO 2014, and MS COCO 2017 datasets also illustrate that SLAM surpasses existing acceleration methods in few-step generation tasks, achieving state-of-the-art performance both on FID and the quality of the generated images.

📄 PDF Abstract BibTeX arXiv:2404.13903

Code (0)

등록된 구현이 없습니다.

Tasks

DenoisingGPUImage GenerationVideo Generation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

PathRelax: Parallel-Path Relaxed Speculative Jacobi Decoding for Accelerating Auto-Regressive Text-to-Image Generation

2026-06-09 · Haodong Lei, Hongsong Wang, Bingxuan Dai, Pan Zhou arxiv

The growing need for high-resolution image generation in autoregressive text-to-image models has resulted in extended token sequences, significantly increasing computational costs and inference times. However, existing s…

Text-to-Image Generation

FastFlow: Accelerating The Generative Flow Matching Models with Bandit Inference

2026-02-11 · Divya Jyoti Bajpai, Dhruv Bhardwaj, Soumya Roy, Tejas Duseja 외 arxiv

Flow-matching models deliver state-of-the-art fidelity in image and video generation, but the inherent sequential denoising process renders them slower. Existing acceleration methods like distillation, trajectory truncat…

Video GenerationImage Generation

Accelerating Power Method with Fast Sketching for Stronger Low-Rank Approximation

2026-05-10 · Shabarish Chenakkod, Michał Dereziński arxiv

The power method is one of the most fundamental tools for extracting top principal components from data through low-rank matrix approximation. Yet, when the target rank is large, the cost of matrix multiplication associa…

MotionFlux: Efficient Text-Guided Motion Generation through Rectified Flow Matching and Preference Alignment

2025-08-27 · Zhiting Gao, Dan Song, Diqiong Jiang, Chao Xue 외 arxiv

Motion generation is essential for animating virtual characters and embodied agents. While recent text-driven methods have made significant strides, they often struggle with achieving precise alignment between linguistic…

Global universality via discrete-time signatures

2026-03-10 · Mihriban Ceylan, David J. Prömel arxiv

We establish global universal approximation theorems for non-anticipative and general path-dependent functionals on spaces of piecewise linear paths, stating that linear functionals of the corresponding signatures are de…