paper-with-me

홈 › Papers

LearniBridge: Learnable Calibration of Feature Caching for Diffusion Models Acceleration

2026-06-25 · Xuyue Huang, Zhe Chen, Wang Shen, Xiao-Ping Zhang arxiv

Diffusion Transformers (DiTs) have driven substantial progress in image and video generation but suffer from prohibitive computational costs. Feature caching accelerates inference by reusing intermediate representations. Existing methods rely on historical features for implementation simplicity, yet suffer from severe error accumulation at high acceleration ratios. To address this limitation, we investigate the nature of the requisite feature correction. We demonstrate that the optimal calibration update is characterized by a shared low-rank subspace across diverse prompts. Guided by this structural insight, we propose LearniBridge, a learnable calibration mechanism for feature caching that bridges multiple timesteps through lightweight LoRA updates. This mechanism enables effective calibration requiring only 3-5 training samples. Extensive experiments on image and video generation show that LearniBridge achieves up to $5.87\times$, $5.75\times$, and $4.10\times$ acceleration on FLUX, HunyuanVideo, and WAN2.1, respectively. On WAN2.1, it improves VBench by 1.28% over the previous SOTA at $4.10\times$ acceleration. Our code is available at https://github.com/Iiiiiiirene/LearniBridge.

📄 PDF Abstract BibTeX arXiv:2606.26778

Code (0)

등록된 구현이 없습니다.

Tasks

Video Generation

Similar Papers 제목 키워드 기반

DisCa: Accelerating Video Diffusion Transformers with Distillation-Compatible Learnable Feature Caching

2026-02-05 · Chang Zou, Changlin Li, Yang Li, Patrol Li 외 arxiv

While diffusion models have achieved great success in the field of video generation, this progress is accompanied by a rapidly escalating computational burden. Among the existing acceleration methods, Feature Caching is …

Video GenerationImage Generation

LinCa: Accelerating Diffusion Models via Learnable Decomposed Feature Caching

2026-08-18 · Jinshan Liu, Haoran Qin, Xiaobing Tu, Jiacheng Liu 외 arxiv

Diffusion models have achieved remarkable success in image and video generation, yet the high computational cost of iterative sampling remains a critical bottleneck for practical deployment. Feature caching has emerged a…

Video Generation

Accelerating Diffusion Transformers with Dual Feature Caching

2024-12-25 · Chang Zou, Evelyn Zhang, Runlin Guo, Haohang Xu 외

Diffusion Transformers (DiT) have become the dominant methods in image and video generation yet still suffer substantial computational costs. As an effective approach for DiT acceleration, feature caching methods are des…

Video Generation

Beyond Fixed Formulas: Data-Driven Linear Predictor for Efficient Diffusion Models

2026-04-29 · Zhirong Shen, Rui Huang, Jiacheng Liu, Chang Zou 외 arxiv

To address the high sampling cost of Diffusion Transformers (DiTs), feature caching offers a training-free acceleration method. However, existing methods rely on hand-crafted forecasting formulas that fail under aggressi…

Evolutionary Caching to Accelerate Your Off-the-Shelf Diffusion Model

2025-06-18 · Anirud Aggarwal, Abhinav Shrivastava, Matthew Gwilliam

Diffusion-based image generation models excel at producing high-quality synthetic content, but suffer from slow and computationally expensive inference. Prior work has attempted to mitigate this by caching and reusing fe…

Image Generation