paper-with-me

홈 › Papers

Zero-Shot Personalization of Objects via Textual Inversion

2026-03-24 · Aniket Roy, Maitreya Suin, Rama Chellappa arxiv

Recent advances in text-to-image diffusion models have substantially improved the quality of image customization, enabling the synthesis of highly realistic images. Despite this progress, achieving fast and efficient personalization remains a key challenge, particularly for real-world applications. Existing approaches primarily accelerate customization for human subjects by injecting identity-specific embeddings into diffusion models, but these strategies do not generalize well to arbitrary object categories, limiting their applicability. To address this limitation, we propose a novel framework that employs a learned network to predict object-specific textual inversion embeddings, which are subsequently integrated into the UNet timesteps of a diffusion model for text-conditional customization. This design enables rapid, zero-shot personalization of a wide range of objects in a single forward pass, offering both flexibility and scalability. Extensive experiments across multiple tasks and settings demonstrate the effectiveness of our approach, highlighting its potential to support fast, versatile, and inclusive image customization. To the best of our knowledge, this work represents the first attempt to achieve such general-purpose, training-free personalization within diffusion models, paving the way for future research in personalized image generation.

📄 PDF Abstract BibTeX arXiv:2603.23010

Code (0)

등록된 구현이 없습니다.

Tasks

Personalized Image Generation

Similar Papers 제목 키워드 기반

Personalization as a Shortcut for Few-Shot Backdoor Attack against Text-to-Image Diffusion Models

2023-05-18 · Yihao Huang, Felix Juefei-Xu, Qing Guo, Jie Zhang 외

Although recent personalization methods have democratized high-resolution image synthesis by enabling swift concept acquisition with minimal examples and lightweight computation, they also present an exploitable avenue f…

Backdoor AttackImage Generation

Specialist Diffusion: Plug-and-Play Sample-Efficient Fine-Tuning of Text-to-Image Diffusion Models To Learn Any Unseen Style

2023-01-01 · CVPR 2023 1 · Haoming Lu, Hazarapet Tunanyan, Kai Wang, Shant Navasardyan 외

Diffusion models have demonstrated impressive capability of text-conditioned image synthesis, and broader application horizons are emerging by personalizing those pretrained diffusion models toward generating some sp…

DisentanglementImage Generation

Zero-Shot Composed Image Retrieval with Textual Inversion

2023-03-27 · ICCV 2023 1 · Alberto Baldrati, Lorenzo Agnolucci, Marco Bertini, Alberto del Bimbo

Composed Image Retrieval (CIR) aims to retrieve a target image based on a query composed of a reference image and a relative caption that describes the difference between the two images. The high effort and cost required…

Composed Image Retrieval (CoIR)Image RetrievalRetrievalZero-Shot Composed Image Retrieval (ZS-CIR)

Zero-shot Composed Image Retrieval Considering Query-target Relationship Leveraging Masked Image-text Pairs

2024-06-27 · Huaying Zhang, rintaro yanagi, Ren Togo, Takahiro Ogawa 외

This paper proposes a novel zero-shot composed image retrieval (CIR) method considering the query-target relationship by masked image-text pairs. The objective of CIR is to retrieve the target image using a query image a…

Image RetrievalLanguage ModelingLanguage ModellingRetrieval

Textual Inversion for Efficient Adaptation of Open-Vocabulary Object Detectors Without Forgetting

2025-08-07 · Frank Ruis, Gertjan Burghouts, Hugo Kuijf arxiv

Recent progress in large pre-trained vision language models (VLMs) has reached state-of-the-art performance on several object detection benchmarks and boasts strong zero-shot capabilities, but for optimal performance on …

Transfer LearningObject Detection