paper-with-me

홈 › Papers

Leveraging Semantic Attribute Binding for Free-Lunch Color Control in Diffusion Models

2025-03-12 · Héctor Laria, Alexandra Gomez-Villa, Jiang Qin, Muhammad Atif Butt, Bogdan Raducanu, Javier Vazquez-Corral, Joost Van de Weijer, Kai Wang

Recent advances in text-to-image (T2I) diffusion models have enabled remarkable control over various attributes, yet precise color specification remains a fundamental challenge. Existing approaches, such as ColorPeel, rely on model personalization, requiring additional optimization and limiting flexibility in specifying arbitrary colors. In this work, we introduce ColorWave, a novel training-free approach that achieves exact RGB-level color control in diffusion models without fine-tuning. By systematically analyzing the cross-attention mechanisms within IP-Adapter, we uncover an implicit binding between textual color descriptors and reference image features. Leveraging this insight, our method rewires these bindings to enforce precise color attribution while preserving the generative capabilities of pretrained models. Our approach maintains generation quality and diversity, outperforming prior methods in accuracy and applicability across diverse object categories. Through extensive evaluations, we demonstrate that ColorWave establishes a new paradigm for structured, color-consistent diffusion-based image synthesis.

📄 PDF Abstract BibTeX arXiv:2503.09864

Code (0)

등록된 구현이 없습니다.

Tasks

AttributeDiversityImage Generation

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

ABE-CLIP: Training-Free Attribute Binding Enhancement for Compositional Image-Text Matching

2025-12-19 · Qi Zhang, Yuxu Chen, Lei Deng, Lili Shen arxiv

Contrastive Language-Image Pretraining (CLIP) has achieved remarkable performance in various multimodal tasks. However, it still struggles with compositional image-text matching, particularly in accurately associating ob…

Image-text matching

Token Merging for Training-Free Semantic Binding in Text-to-Image Synthesis

2024-11-11 · Taihang Hu, Linxuan Li, Joost Van de Weijer, Hongcheng Gao 외

Although text-to-image (T2I) models exhibit remarkable generation capabilities, they frequently fail to accurately bind semantically related objects or attributes in the input prompts; a challenge termed semantic binding…

AttributeImage GenerationObject

How Bias Binds: Measuring Hidden Associations for Bias Control in Text-to-Image Compositions

2025-11-10 · Jeng-Lin Li, Ming-Ching Chang, Wei-Chao Chen arxiv

Text-to-image generative models often exhibit bias related to sensitive attributes. However, current research tends to focus narrowly on single-object prompts with limited contextual diversity. In reality, each object or…

On Geometrical Properties of Text Token Embeddings for Strong Semantic Binding in Text-to-Image Generation

2025-03-29 · Hoigi Seo, Junseo Bang, Haechang Lee, Joohoon Lee 외

Text-to-Image (T2I) models often suffer from text-image misalignment in complex scenes involving multiple objects and attributes. Semantic binding aims to mitigate this issue by accurately associating the generated attri…

Image GenerationText to Image GenerationText-to-Image Generation

Box It to Bind It: Unified Layout Control and Attribute Binding in T2I Diffusion Models

2024-02-27 · Ashkan Taghipour, Morteza Ghahremani, Mohammed Bennamoun, Aref Miri Rekavandi 외

While latent diffusion models (LDMs) excel at creating imaginative images, they often lack precision in semantic fidelity and spatial control over where objects are generated. To address these deficiencies, we introduce …

Attribute