paper-with-me

Papers

MultiCompose: Multi-Concept Personalized Composition with Per-Subject Attribute Binding

2026-08-04 · Ruirui Zhang, Zhengkai Zhao, Pan Gao arxiv

Text-to-image diffusion models enable personalization of specific visual concepts from a small number of reference images. However, generating a single image that contains multiple personalized subjects, each bound to user-specified attributes such as clothing, accessories, and held objects, remains largely unaddressed. Without explicit spatial constraints, concurrently activated concept checkpoints produce overlapping cross-attention responses, causing per-subject identity degradation and attribute misalignment. Moreover, no established benchmark jointly evaluates these two failure modes in the personalized multi-subject setting. We present MultiCompose, a composition framework that decouples per-concept personalization from multi-subject inference. A semantic preservation regularization maintains attribute binding capacity during fine-tuning, while a two-phase inference procedure automatically establishes subject layout and composes per-concept predictions through spatially exclusive masks. We further introduce MSP-Bench, a benchmark that jointly evaluates identity fidelity (ID), attribute binding accuracy (BIND), and attribute misalignment (MIS) through a dual-pathway protocol. Experiments show that MultiCompose outperforms existing methods on both conventional metrics and MSP-Bench, confirming the benchmark's ability to reveal failure modes that conventional metrics overlook. Code is available at https://github.com/I2-Multimedia-Lab/MultiCompose

📄 PDF Abstract BibTeX arXiv:2608.03708

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Photoswap: Personalized Subject Swapping in Images

2023-05-29 · NeurIPS 2023 11

In an era where images and visual content dominate our digital landscape, the ability to manipulate and personalize these images has become a necessity. Envision seamlessly substituting a tabby cat lounging on a sunlit w…

FreeTuner: Any Subject in Any Style with Training-free Diffusion

2024-05-23 · Youcan Xu, Zhen Wang, Jun Xiao, Wei Liu 외

With the advance of diffusion models, various personalized image generation methods have been proposed. However, almost all existing work only focuses on either subject-driven or style-driven personalization. Meanwhile, …

DisentanglementImage GenerationPersonalized Image Generation

LoRAShop: Training-Free Multi-Concept Image Generation and Editing with Rectified Flow Transformers

2025-05-29 · Yusuf Dalva, Hidir Yesiltepe, Pinar Yanardag

We introduce LoRAShop, the first framework for multi-concept image editing with LoRA models. LoRAShop builds on a key observation about the feature interaction patterns inside Flux-style diffusion transformers: concept-s…

DenoisingImage GenerationVisual Storytelling

LayerComposer: Multi-Human Personalized Generation via Layered Canvas

2025-10-23 · Guocheng Gordon Qian, Ruihang Zhang, Tsai-Shien Chen, Yusuf Dalva 외 arxiv

Despite their impressive visual fidelity, existing personalized image generators lack interactive control over spatial composition and scale poorly to multiple humans. To address these limitations, we present LayerCompos…

Personalized Image Generation

$λ$-ECLIPSE: Multi-Concept Personalized Text-to-Image Diffusion Models by Leveraging CLIP Latent Space

2024-02-07 · Maitreya Patel, Sangmin Jung, Chitta Baral, Yezhou Yang

Despite the recent advances in personalized text-to-image (P-T2I) generative models, it remains challenging to perform finetuning-free multi-subject-driven T2I in a resource-efficient manner. Predominantly, contemporary …

Concept AlignmentGPUPhilosophy