paper-with-me

홈 › Papers

LatexBlend: Scaling Multi-concept Customized Generation with Latent Textual Blending

2025-03-10 · CVPR 2025 1 · Jian Jin, Zhenbo Yu, Yang shen, ZhenYong Fu, Jian Yang

Customized text-to-image generation renders user-specified concepts into novel contexts based on textual prompts. Scaling the number of concepts in customized generation meets a broader demand for user creation, whereas existing methods face challenges with generation quality and computational efficiency. In this paper, we propose LaTexBlend, a novel framework for effectively and efficiently scaling multi-concept customized generation. The core idea of LaTexBlend is to represent single concepts and blend multiple concepts within a Latent Textual space, which is positioned after the text encoder and a linear projection. LaTexBlend customizes each concept individually, storing them in a concept bank with a compact representation of latent textual features that captures sufficient concept information to ensure high fidelity. At inference, concepts from the bank can be freely and seamlessly combined in the latent textual space, offering two key merits for multi-concept generation: 1) excellent scalability, and 2) significant reduction of denoising deviation, preserving coherent layouts. Extensive experiments demonstrate that LaTexBlend can flexibly integrate multiple customized concepts with harmonious structures and high subject fidelity, substantially outperforming baselines in both generation quality and computational efficiency. Our code will be publicly available.

📄 PDF Abstract BibTeX arXiv:2503.06956

Code (0)

등록된 구현이 없습니다.

Tasks

Computational EfficiencyDenoisingImage GenerationText to Image GenerationText-to-Image Generation

Similar Papers 제목 키워드 기반

Non-confusing Generation of Customized Concepts in Diffusion Models

2024-05-11 · Wang Lin, Jingyuan Chen, Jiaxin Shi, Yichen Zhu 외

We tackle the common challenge of inter-concept visual confusion in compositional concept generation using text-guided diffusion models (TGDMs). It becomes even more pronounced in the generation of customized concepts, d…

Infusion: Preventing Customized Text-to-Image Diffusion from Overfitting

2024-04-22 · Weili Zeng, Yichao Yan, Qi Zhu, Zhuo Chen 외

Text-to-image (T2I) customization aims to create images that embody specific visual concepts delineated in textual descriptions. However, existing works still face a main challenge, concept overfitting. To tackle this ch…

MC^2: Multi-concept Guidance for Customized Multi-concept Generation

2025-01-01 · CVPR 2025 1 · Jiaxiu Jiang, Yabo Zhang, Kailai Feng, Xiaohe Wu 외

Customized text-to-image generation, which synthesizes images based on user-specified concepts, has made significant progress in handling individual concepts. However, when extended to multiple concepts, existing met…

Image GenerationText to Image GenerationText-to-Image Generation

FreeCustom: Tuning-Free Customized Image Generation for Multi-Concept Composition

2024-05-22 · CVPR 2024 1 · Ganggui Ding, Canyu Zhao, Wen Wang, Zhen Yang 외

Benefiting from large-scale pre-trained text-to-image (T2I) generative models, impressive progress has been achieved in customized image generation, which aims to generate user-specified concepts. Existing approaches hav…

Image Generation

MC$^2$: Multi-concept Guidance for Customized Multi-concept Generation

2024-04-08 · Jiaxiu Jiang, Yabo Zhang, Kailai Feng, Xiaohe Wu 외

Customized text-to-image generation, which synthesizes images based on user-specified concepts, has made significant progress in handling individual concepts. However, when extended to multiple concepts, existing methods…

Image GenerationText to Image GenerationText-to-Image Generation