paper-with-me

Papers

CustomNet: Zero-shot Object Customization with Variable-Viewpoints in Text-to-Image Diffusion Models

2023-10-30 · Ziyang Yuan, Mingdeng Cao, Xintao Wang, Zhongang Qi, Chun Yuan, Ying Shan

Incorporating a customized object into image generation presents an attractive feature in text-to-image generation. However, existing optimization-based and encoder-based methods are hindered by drawbacks such as time-consuming optimization, insufficient identity preservation, and a prevalent copy-pasting effect. To overcome these limitations, we introduce CustomNet, a novel object customization approach that explicitly incorporates 3D novel view synthesis capabilities into the object customization process. This integration facilitates the adjustment of spatial position relationships and viewpoints, yielding diverse outputs while effectively preserving object identity. Moreover, we introduce delicate designs to enable location control and flexible background control through textual descriptions or specific user-defined images, overcoming the limitations of existing 3D novel view synthesis methods. We further leverage a dataset construction pipeline that can better handle real-world objects and complex backgrounds. Equipped with these designs, our method facilitates zero-shot object customization without test-time optimization, offering simultaneous control over the viewpoints, location, and background. As a result, our CustomNet ensures enhanced identity preservation and generates diverse, harmonious outputs.

📄 PDF Abstract BibTeX arXiv:2310.19784

Code (0)

등록된 구현이 없습니다.

Tasks

Image GenerationNovel View SynthesisObjectText to Image GenerationText-to-Image Generation

Similar Papers 제목 키워드 기반

MS-CustomNet: Controllable Multi-Subject Customization with Hierarchical Relational Semantics

2026-03-22 · Pengxiang Cai, Mengyang Li arxiv

Diffusion-based text-to-image generation has advanced significantly, yet customizing scenes with multiple distinct subjects while maintaining fine-grained control over their interactions remains challenging. Existing met…

Text-to-Image Generation

CustAny: Customizing Anything from A Single Example

2024-06-17 · CVPR 2025 1 · Lingjie Kong, Kai Wu, Xiaobin Hu, Wenhui Han 외

Recent advances in diffusion-based text-to-image models have simplified creating high-fidelity images, but preserving the identity (ID) of specific elements, like a personal dog, is still challenging. Object customizatio…

ObjectVirtual Try-on

HomeDiffusion: Zero-Shot Object Customization with Multi-View Representation Learning for Indoor Scenes

2026-06-29 · Guoqiu Li, Jin Song, Yiyun Fei arxiv

Recently, zero-shot object customization generation methods have rapidly developed and shown tremendous potential for applications. For instance, in the e-commerce domain, consumers can observe the visual effect of furni…

Representation Learning

Z-Magic: Zero-shot Multiple Attributes Guided Image Creator

2025-01-01 · CVPR 2025 1 · Yingying Deng, Xiangyu He, Fan Tang, WeiMing Dong

The customization of multiple attributes has gained increasing popularity with the rising demand for personalized content creation. Despite promising empirical results, the contextual coherence between different attr…

AttributeImage GenerationMulti-Task Learning

Zero-Shot Personalization of Objects via Textual Inversion

2026-03-24 · Aniket Roy, Maitreya Suin, Rama Chellappa arxiv

Recent advances in text-to-image diffusion models have substantially improved the quality of image customization, enabling the synthesis of highly realistic images. Despite this progress, achieving fast and efficient per…

Personalized Image Generation