paper-with-me

Papers

Encoder-based Domain Tuning for Fast Personalization of Text-to-Image Models

2023-02-23 · Rinon Gal, Moab Arar, Yuval Atzmon, Amit H. Bermano, Gal Chechik, Daniel Cohen-Or

Text-to-image personalization aims to teach a pre-trained diffusion model to reason about novel, user provided concepts, embedding them into new scenes guided by natural language prompts. However, current personalization approaches struggle with lengthy training times, high storage requirements or loss of identity. To overcome these limitations, we propose an encoder-based domain-tuning approach. Our key insight is that by underfitting on a large set of concepts from a given domain, we can improve generalization and create a model that is more amenable to quickly adding novel concepts from the same domain. Specifically, we employ two components: First, an encoder that takes as an input a single image of a target concept from a given domain, e.g. a specific face, and learns to map it into a word-embedding representing the concept. Second, a set of regularized weight-offsets for the text-to-image model that learn how to effectively ingest additional concepts. Together, these components are used to guide the learning of unseen concepts, allowing us to personalize a model using only a single image and as few as 5 training steps - accelerating personalization from dozens of minutes to seconds, while preserving quality.

📄 PDF Abstract BibTeX arXiv:2302.12228

Code (0)

등록된 구현이 없습니다.

Tasks

Novel Concepts

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Domain-Agnostic Tuning-Encoder for Fast Personalization of Text-To-Image Models

2023-07-13 · Moab Arar, Rinon Gal, Yuval Atzmon, Gal Chechik 외

Text-to-image (T2I) personalization allows users to guide the creative image generation process by combining their own visual concepts in natural language prompts. Recently, encoder-based techniques have emerged as a new…

Image Generation

LCM-Lookahead for Encoder-based Text-to-Image Personalization

2024-04-04 · Rinon Gal, Or Lichter, Elad Richardson, Or Patashnik 외

Recent advancements in diffusion models have introduced fast sampling methods that can effectively produce high-quality images in just one or a few denoising steps. Interestingly, when these are distilled from existing d…

DenoisingDiversity

JeDi: Joint-Image Diffusion Models for Finetuning-Free Personalized Text-to-Image Generation

2024-07-08 · CVPR 2024 1 · Yu Zeng, Vishal M. Patel, Haochen Wang, Xun Huang 외

Personalized text-to-image generation models enable users to create images that depict their individual possessions in diverse scenes, finding applications in various domains. To achieve the personalization capability, e…

Dataset GenerationImage GenerationText to Image GenerationText-to-Image Generation

InstantBooth: Personalized Text-to-Image Generation without Test-Time Finetuning

2023-04-06 · CVPR 2024 1 · Jing Shi, Wei Xiong, Zhe Lin, Hyun Joon Jung

Recent advances in personalized image generation allow a pre-trained text-to-image model to learn a new concept from a set of images. However, existing personalization approaches usually require heavy test-time finetunin…

Diffusion PersonalizationDiffusion Personalization Tuning FreeImage GenerationPersonalized Image Generation+2

HyperDreamBooth: HyperNetworks for Fast Personalization of Text-to-Image Models

2023-07-13 · CVPR 2024 1 · Nataniel Ruiz, Yuanzhen Li, Varun Jampani, Wei Wei 외

Personalization has emerged as a prominent aspect within the field of generative AI, enabling the synthesis of individuals in diverse contexts and styles, while retaining high-fidelity to their identities. However, the p…

Diffusion PersonalizationDiffusion Personalization Tuning FreeDiversityGPU