paper-with-me

홈 › Papers

When Few Steps Are Enough: Training-Free Acceleration of Identity-Preserved Generation

2026-05-10 · Dongqi Zheng arxiv

Identity-preserved image generation is typically built on many-step diffusion backbones, making personalized generation expensive at deployment time. We show that this cost is often unnecessary for identity-conditioned FLUX generation. A frozen InfuseNet identity adapter trained with dev transfers directly to the distilled schnell backbone without retraining. This two-line replacement -- changing the backbone path and disabling classifier-free guidance -- reduces latency by 5.9x while improving ArcFace identity similarity by +0.028 and lpips by -0.016 over the standard 28-step dev baseline. To explain why this works, we analyze the denoising trajectory and find that identity fidelity enters an early effective regime, often within 4-8 steps, while later steps primarily refine visual detail, sharpness, and contrast. Adapter ablations confirm that identity formation depends on the identity adapter, while attention-stream norm probes suggest that the relative conditioning contribution decreases as sampling proceeds. Preliminary style-adapter and object-adapter sweeps on SDXL and SD1.5 show similar diminishing returns after intermediate steps. These results position distilled backbone replacement as a simple, training-free strategy for improving the efficiency-fidelity tradeoff of identity-preserved generation.

📄 PDF Abstract BibTeX arXiv:2605.09460

Code (0)

등록된 구현이 없습니다.

Tasks

Image Generation

Similar Papers 제목 키워드 기반

Denoising as Path Planning: Training-Free Acceleration of Diffusion Models with DPCache

2026-02-26 · Bowen Cui, Yuanbin Wang, Huajiang Xu, Biaolong Chen 외 arxiv

Diffusion models have demonstrated remarkable success in image and video generation, yet their practical deployment remains hindered by the substantial computational overhead of multi-step iterative sampling. Among accel…

Video Generation

Training Acceleration of Low-Rank Decomposed Networks using Sequential Freezing and Rank Quantization

2023-09-07 · Habib Hajimolahoseini, Walid Ahmed, Yang Liu

Low Rank Decomposition (LRD) is a model compression technique applied to the weight tensors of deep learning models in order to reduce the number of trainable parameters and computational complexity. However, due to high…

Model CompressionQuantization

Light Interaction: Training-Free Inference Acceleration for Interactive Video World Models

2026-05-29 · Jiacheng Lu, Haoyi Zhu, Sipei Yi, Enze Xie 외 arxiv

Interactive video world models generate video chunk by chunk in response to user-controlled camera movements, enabling applications such as real-time game simulation, virtual scene navigation, and embodied AI training. H…

DisCa: Accelerating Video Diffusion Transformers with Distillation-Compatible Learnable Feature Caching

2026-02-05 · Chang Zou, Changlin Li, Yang Li, Patrol Li 외 arxiv

While diffusion models have achieved great success in the field of video generation, this progress is accompanied by a rapidly escalating computational burden. Among the existing acceleration methods, Feature Caching is …

Video GenerationImage Generation

Predictive Feature Caching for Training-free Acceleration of Molecular Geometry Generation

2025-10-06 · Johanna Sommer, John Rachwan, Nils Fleischmann, Stephan Günnemann 외 arxiv

Flow matching models generate high-fidelity molecular geometries but incur significant computational costs during inference, requiring hundreds of network evaluations. This inference overhead becomes the primary bottlene…