paper-with-me

홈 › Papers

HarmoniCa: Harmonizing Training and Inference for Better Feature Caching in Diffusion Transformer Acceleration

2024-10-02 · Yushi Huang, Zining Wang, Ruihao Gong, Jing Liu, Xinjie Zhang, Jinyang Guo, Xianglong Liu, Jun Zhang

Diffusion Transformers (DiTs) excel in generative tasks but face practical deployment challenges due to high inference costs. Feature caching, which stores and retrieves redundant computations, offers the potential for acceleration. Existing learning-based caching, though adaptive, overlooks the impact of the prior timestep. It also suffers from misaligned objectives--aligned predicted noise vs. high-quality images--between training and inference. These two discrepancies compromise both performance and efficiency. To this end, we harmonize training and inference with a novel learning-based caching framework dubbed HarmoniCa. It first incorporates Step-Wise Denoising Training (SDT) to ensure the continuity of the denoising process, where prior steps can be leveraged. In addition, an Image Error Proxy-Guided Objective (IEPO) is applied to balance image quality against cache utilization through an efficient proxy to approximate the image error. Extensive experiments across $8$ models, $4$ samplers, and resolutions from $256\times256$ to $2K$ demonstrate superior performance and speedup of our framework. For instance, it achieves over $40\%$ latency reduction (i.e., $2.07\times$ theoretical speedup) and improved performance on PixArt-$\alpha$. Remarkably, our image-free approach reduces training time by $25\%$ compared with the previous method. Our code is available at https://github.com/ModelTC/HarmoniCa.

📄 PDF Abstract BibTeX arXiv:2410.01723

Code (1)

modeltc/harmonica 공식 구현 pytorch

Tasks

2kDenoising

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Olivia: Harmonizing Time Series Foundation Models with Power Spectral Density

2026-05-17 · Jingru Fei, Kun Yi, Alex Xing Wang, Qingsong Wen 외 arxiv

Time series foundation models rely on large-scale pretraining over diverse datasets across domains, yet their heterogeneity in temporal patterns could hinder the effectiveness of training and learning transferable time s…

HarmonicAttack: An Adaptive Cross-Domain Audio Watermark Removal

2025-11-26 · Kexin Li, Xiao Hu, Ilya Grishchenko, David Lie arxiv

The availability of high-quality, AI-generated audio raises security challenges such as misinformation campaigns and voice-cloning fraud. A key defense against the misuse of AI-generated audio is by watermarking it, so t…

MaxEnt Learners are Biased Against Giving Probability to Harmonically Bounded Candidates

2022-02-01 · SCiL 2022 2 · Charlie O’Hara

Harmonica: A Self-Adaptation Exemplar for Sustainable MLOps

2026-01-17 · Ananya Halgatti, Shaunak Biswas, Hiya Bhatt, Srinivasan Rakhunathan 외 arxiv

Machine learning enabled systems (MLS) often operate in settings where they regularly encounter uncertainties arising from changes in their surrounding environment. Without structured oversight, such changes can degrade …

Time Series Regression

HarmonICA: Neural non-stationarity correction and source separation for motor neuron interfaces

2024-06-28 · Alexander Kenneth Clarke, Agnese Grison, Irene Mendez Guerra, Pranav Mamidanna 외

A major outstanding problem when interfacing with spinal motor neurons is how to accurately compensate for non-stationary effects in the signal during source separation routines, particularly when they cannot be estimate…