paper-with-me

홈 › Papers

Orthogonal Negative Guidance in Attention Feature Space for Text-to-Image Generation

2026-05-28 · Jungmin Ko, Jungwon Park, Jimyeong Kim, Changin Choi, Wonseok Lee, Wonjong Rhee arxiv

Text-to-image (T2I) models have become increasingly capable of generating high-quality images. Yet, enforcing the explicit absence of a specified object or attribute remains a fundamentally challenging problem. Existing approaches, including prompt negation, post-hoc editing, and negative guidance, remain insufficient for explicit concept suppression, often failing to remove the target concept or degrading overall image quality. To this end, we propose Orthogonal Negative Guidance in attention feature space, a training-free method that operates in the attention output space of MM-DiT-based T2I transformers. Our method orthogonalizes negative-prompt attention features with respect to positive-prompt features and subtracts only the orthogonal component, suppressing unwanted concepts while preserving desired semantics. Experiments on FLUX-dev and FLUX-schnell show that our method achieves favorable trade-offs between concept suppression, prompt alignment, and image quality. In human evaluation, our method outperforms the second-best baseline by 18.78%. We further show that our method supports multi-concept suppression and adjustable concept suppression.

📄 PDF Abstract BibTeX arXiv:2605.29390

Code (0)

등록된 구현이 없습니다.

Tasks

Text-to-Image Generation

Similar Papers 제목 키워드 기반

OrthoTryOn: Geometric Orthogonalization for Conflict-Free Unified Fashion Generation

2026-06-26 · Zhaotong Yang, Ying Tai, Jiahui Zhan, Yu Zheng 외 arxiv

Unified fashion generation integrates tasks like virtual try-on and garment reconstruction into a single model to reduce task-specific adaptation costs. However, naive parameter sharing across semantically distinct tasks…

Garment ReconstructionVirtual Try-on

EMAG: Self-Rectifying Diffusion Sampling with Exponential Moving Average Guidance

2025-12-19 · Ankit Yadav, Ta Duc Huy, Lingqiao Liu arxiv

In diffusion and flow-matching generative models, guidance techniques are widely used to improve sample quality and consistency. Classifier-free guidance (CFG) is the de facto choice in modern systems and achieves this b…

Dual Orthogonal Guidance for Robust Diffusion-based Handwritten Text Generation

2025-08-23 · Konstantina Nikolaidou, George Retsinas, Giorgos Sfikas, Silvia Cascianelli 외 arxiv

Diffusion-based Handwritten Text Generation (HTG) approaches achieve impressive results on frequent, in-vocabulary words observed at training time and on regular styles. However, they are prone to memorizing training sam…

Text Generation

OrthoPhys: Physically Plausible Video Generation with Orthogonal-View Geometry Guidance

2026-03-19 · Cong Wang, Hanxin Zhu, Xiao Tang, Jiayi Luo 외 arxiv

Recent progress in video generation has led to substantial improvements in visual fidelity, yet ensuring physically consistent motion remains a fundamental challenge. Intuitively, this limitation can be attributed to the…

Video Generation

Geometry-Aware Attention Guidance for Diffusion Models via Modern Hopfield Dynamics

2026-03-03 · Kwanyoung Kim arxiv

Classifier-Free Guidance (CFG) improves sample quality in diffusion models, but its dual-pass inference and reliance on null-condition training limit its use in few-step regimes. Attention-space guidance has emerged as a…