paper-with-me

홈 › Papers

DiffuseKronA: A Parameter Efficient Fine-tuning Method for Personalized Diffusion Models

2024-02-27 · Shyam Marjit, Harshit Singh, Nityanand Mathur, Sayak Paul, Chia-Mu Yu, Pin-Yu Chen

In the realm of subject-driven text-to-image (T2I) generative models, recent developments like DreamBooth and BLIP-Diffusion have led to impressive results yet encounter limitations due to their intensive fine-tuning demands and substantial parameter requirements. While the low-rank adaptation (LoRA) module within DreamBooth offers a reduction in trainable parameters, it introduces a pronounced sensitivity to hyperparameters, leading to a compromise between parameter efficiency and the quality of T2I personalized image synthesis. Addressing these constraints, we introduce \textbf{\textit{DiffuseKronA}}, a novel Kronecker product-based adaptation module that not only significantly reduces the parameter count by 35\% and 99.947\% compared to LoRA-DreamBooth and the original DreamBooth, respectively, but also enhances the quality of image synthesis. Crucially, \textit{DiffuseKronA} mitigates the issue of hyperparameter sensitivity, delivering consistent high-quality generations across a wide range of hyperparameters, thereby diminishing the necessity for extensive fine-tuning. Furthermore, a more controllable decomposition makes \textit{DiffuseKronA} more interpretable and even can achieve up to a 50\% reduction with results comparable to LoRA-Dreambooth. Evaluated against diverse and complex input images and text prompts, \textit{DiffuseKronA} consistently outperforms existing models, producing diverse images of higher quality with improved fidelity and a more accurate color distribution of objects, all the while upholding exceptional parameter efficiency, thus presenting a substantial advancement in the field of T2I generative modeling. Our project page, consisting of links to the code, and pre-trained checkpoints, is available at https://diffusekrona.github.io/.

📄 PDF Abstract BibTeX arXiv:2402.17412

Code (0)

등록된 구현이 없습니다.

Tasks

Image Generationparameter-efficient fine-tuningSensitivity

Similar Papers 제목 키워드 기반

ViCo: Plug-and-play Visual Condition for Personalized Text-to-image Generation

2023-06-01 · Shaozhe Hao, Kai Han, Shihao Zhao, Kwan-Yee K. Wong

Personalized text-to-image generation using diffusion models has recently emerged and garnered significant interest. This task learns a novel concept (e.g., a unique toy), illustrated in a handful of images, into a gener…

Image GenerationText to Image GenerationText-to-Image Generation

HiFi Tuner: High-Fidelity Subject-Driven Fine-Tuning for Diffusion Models

2023-11-30 · Zhonghao Wang, Wei Wei, Yang Zhao, Zhisheng Xiao 외

This paper explores advancements in high-fidelity personalized image generation through the utilization of pre-trained text-to-image diffusion models. While previous approaches have made significant strides in generating…

DenoisingImage Generationparameter-efficient fine-tuningPersonalized Image Generation

Subject-Diffusion:Open Domain Personalized Text-to-Image Generation without Test-time Fine-tuning

2023-07-21 · Jian Ma, Junhao Liang, Chen Chen, Haonan Lu

Recent progress in personalized image generation using diffusion models has been significant. However, development in the area of open-domain and non-fine-tuning personalized image generation is proceeding rather slowly.…

Diffusion PersonalizationDiffusion Personalization Tuning FreeImage GenerationPersonalized Image Generation+2

Exploring Sparsity for Parameter Efficient Fine Tuning Using Wavelets

2025-05-18 · Ahmet Bilican, M. Akin Yilmaz, A. Murat Tekalp, R. Gökberk Cinbiş

Efficiently adapting large foundation models is critical, especially with tight compute and memory budgets. Parameter-Efficient Fine-Tuning (PEFT) methods such as LoRA offer limited granularity and effectiveness in few-p…

DiversityImage Generationparameter-efficient fine-tuningText to Image Generation+1

Conceptwm: A Diffusion Model Watermark for Concept Protection

2024-11-18 · Liangqi Lei, Keke Gai, Jing Yu, Liehuang Zhu 외

The personalization techniques of diffusion models succeed in generating specific concepts but also pose threats to copyright protection and illegal use. Model Watermarking is an effective method to prevent the unauthori…

Image Generation