paper-with-me

Papers

Diffusion Cocktail: Mixing Domain-Specific Diffusion Models for Diversified Image Generations

2023-12-12 · Haoming Liu, Yuanhe Guo, Shengjie Wang, Hongyi Wen

Diffusion models, capable of high-quality image generation, receive unparalleled popularity for their ease of extension. Active users have created a massive collection of domain-specific diffusion models by fine-tuning base models on self-collected datasets. Recent work has focused on improving a single diffusion model by uncovering semantic and visual information encoded in various architecture components. However, those methods overlook the vastly available set of fine-tuned diffusion models and, therefore, miss the opportunity to utilize their combined capacity for novel generation. In this work, we propose Diffusion Cocktail (Ditail), a training-free method that transfers style and content information between multiple diffusion models. This allows us to perform diversified generations using a set of diffusion models, resulting in novel images unobtainable by a single model. Ditail also offers fine-grained control of the generation process, which enables flexible manipulations of styles and contents. With these properties, Ditail excels in numerous applications, including style transfer guided by diffusion models, novel-style image generation, and image manipulation via prompts or collage inputs.

📄 PDF Abstract BibTeX arXiv:2312.08873

Code (1)

MAPS-research/Ditail 공식 구현 pytorch

Tasks

Image GenerationImage ManipulationStyle Transfer

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically
BASE 설명 없음
Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Cocktail: Mixing Multi-Modality Controls for Text-Conditional Image Generation

2023-06-01 · Minghui Hu, Jianbin Zheng, Daqing Liu, Chuanxia Zheng 외

Text-conditional diffusion models are able to generate high-fidelity images with diverse contents. However, linguistic representations frequently exhibit ambiguous descriptions of the envisioned objective imagery, requir…

Conditional Image GenerationImage Generation

Cocktail: Mixing Multi-Modality Control for Text-Conditional Image Generation

2023-09-21 · NeurIPS 2023 11

Text-conditional diffusion models are able to generate high-fidelity images with diverse contents. However, linguistic representations frequently exhibit ambiguous descriptions of the envisioned objective imagery, requir…

Separate And Diffuse: Using a Pretrained Diffusion Model for Improving Source Separation

2023-01-25 · Shahar Lutati, Eliya Nachmani, Lior Wolf

The problem of speech separation, also known as the cocktail party problem, refers to the task of isolating a single speech signal from a mixture of speech signals. Previous work on source separation derived an upper bou…

Audio Source SeparationGeneralization BoundsMulti-Speaker Source SeparationSpeech Separation

Prompt Mixing in Diffusion Models using the Black Scholes Algorithm

2024-05-22 · Divya Kothandaraman, Ming Lin, Dinesh Manocha

We introduce a novel approach for prompt mixing, aiming to generate images at the intersection of multiple text prompts using pre-trained text-to-image diffusion models. At each time step during diffusion denoising, our …

Denoising

Diffusion-Guided Mask-Consistent Paired Mixing for Endoscopic Image Segmentation

2025-11-05 · Pengyu Jie, Wanquan Liu, Rui He, Yihui Wen 외 arxiv

Augmentation for dense prediction typically relies on either sample mixing or generative synthesis. Mixing improves robustness but misaligned masks yield soft label ambiguity. Diffusion synthesis increases apparent diver…

Image Segmentation