paper-with-me

홈 › Papers

Beyond Fixed Anchors: Precisely Erasing Concepts with Sibling Exclusive Counterparts

2025-10-18 · Tong Zhang, Ru Zhang, Jianyi Liu, Zhen Yang, Gongshen Liu arxiv

Existing concept erasure methods for text-to-image diffusion models commonly rely on fixed anchor strategies, which often lead to critical issues such as concept re-emergence and erosion. To address this, we conduct causal tracing to reveal the inherent sensitivity of erasure to anchor selection and define Sibling Exclusive Concepts as a superior class of anchors. Based on this insight, we propose \textbf{SELECT} (Sibling-Exclusive Evaluation for Contextual Targeting), a dynamic anchor selection framework designed to overcome the limitations of fixed anchors. Our framework introduces a novel two-stage evaluation mechanism that automatically discovers optimal anchors for precise erasure while identifying critical boundary anchors to preserve related concepts. Extensive evaluations demonstrate that SELECT, as a universal anchor solution, not only efficiently adapts to multiple erasure frameworks but also consistently outperforms existing baselines across key performance metrics, averaging only 4 seconds for anchor mining of a single concept.

📄 PDF Abstract BibTeX arXiv:2510.16342

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Beyond Text Prompts: Precise Concept Erasure through Text-Image Collaboration

2026-04-17 · Jun Li, Lizhi Xiong, Ziqiang Li, Weiwei Jiang 외 arxiv

Text-to-image generative models have achieved impressive fidelity and diversity, but can inadvertently produce unsafe or undesirable content due to implicit biases embedded in large-scale training datasets. Existing conc…

Text-to-Image GenerationRepresentation Learning

Fantastic Targets for Concept Erasure in Diffusion Models and Where To Find Them

2025-01-31 · Anh Bui, Trang Vu, Long Vuong, Trung Le 외

Concept erasure has emerged as a promising technique for mitigating the risk of harmful content generation in diffusion models by selectively unlearning undesirable concepts. The common principle of previous works to rem…

RealEra: Semantic-level Concept Erasure via Neighbor-Concept Mining

2024-10-11 · Yufan Liu, Jinyang An, Wanqian Zhang, Ming Li 외

The remarkable development of text-to-image generation models has raised notable security concerns, such as the infringement of portrait rights and the generation of inappropriate content. Concept erasure has been propos…

Image GenerationSpecificityText to Image GenerationText-to-Image Generation

SAGE: Exploring the Boundaries of Unsafe Concept Domain with Semantic-Augment Erasing

2025-06-11 · Hongguang Zhu, Yunchao Wei, Mengyu Wang, Siyu Jiao 외

Diffusion models (DMs) have achieved significant progress in text-to-image generation. However, the inevitable inclusion of sensitive information during pre-training poses safety risks, such as unsafe content generation …

Image GenerationText to Image GenerationText-to-Image Generation

IdentityGuard: Context-Aware Restriction and Provenance for Personalized Synthesis

2026-03-14 · Lingyun Zhang, Yu Xie, Ping Chen arxiv

The nature of personalized text-to-image models poses a unique safety challenge that generic context-blind methods are ill-equipped to handle. Such global filters create a dilemma: to prevent misuse, they are forced to d…