paper-with-me

홈 › Papers

Now You See It, Now You Don't - Instant Concept Erasure for Safe Text-to-Image and Video Generation

2025-11-24 · Shristi Das Biswas, Arani Roy, Kaushik Roy arxiv

Robust concept removal for text-to-image (T2I) and text-to-video (T2V) models is essential for their safe deployment. Existing methods, however, suffer from costly retraining, inference overhead, or vulnerability to adversarial attacks. Crucially, they rarely model the latent semantic overlap between the target erase concept and surrounding content -- causing collateral damage post-erasure -- and even fewer methods work reliably across both T2I and T2V domains. We introduce Instant Concept Erasure (ICE), a training-free, modality-agnostic, one-shot weight modification approach that achieves precise, persistent unlearning with zero overhead. ICE defines erase and preserve subspaces using anisotropic energy-weighted scaling, then explicitly regularises against their intersection using a unique, closed-form overlap projector. We pose a convex and Lipschitz-bounded Spectral Unlearning Objective, balancing erasure fidelity and intersection preservation, that admits a stable and unique analytical solution. This solution defines a dissociation operator that is translated to the model's text-conditioning layers, making the edit permanent and runtime-free. Across targeted removals of artistic styles, objects, identities, and explicit content, ICE efficiently achieves strong erasure with improved robustness to red-teaming, all while causing only minimal degradation of original generative abilities in both T2I and T2V models.

📄 PDF Abstract BibTeX arXiv:2511.18684

Code (0)

등록된 구현이 없습니다.

Tasks

Video Generation

Similar Papers 제목 키워드 기반

ScaleErasure: Inference-Time Minimal Intervention for Precise Concept Erasure in Next-Scale Autoregressive Image Generation

2026-06-28 · Cong Wang, Haiyu Wu, Zhiwei Jiang, Zifeng Cheng 외 arxiv

Concept erasure aims to prevent image generative models from producing unsafe content while preserving their general generative capability. Meanwhile, next-scale autoregressive (AR) image generation has recently emerged …

Image Generation

Closing the Safety Gap: Surgical Concept Erasure in Visual Autoregressive Models

2025-09-26 · Xinhao Zhong, Yimin Zhou, Zhiqi Zhang, Junhao Li 외 arxiv

The rapid progress of visual autoregressive (VAR) models has brought new opportunities for text-to-image generation, but also heightened safety concerns. Existing concept erasure techniques, primarily designed for diffus…

Text-to-Image Generation

Bi-Erasing: A Bidirectional Framework for Concept Removal in Diffusion Models

2025-12-15 · Hao Chen, Yiwei Wang, Songze Li arxiv

Concept erasure, which fine-tunes diffusion models to remove undesired or harmful visual concepts, has become a mainstream approach to mitigating unsafe or illegal image generation in text-to-image models.However, existi…

Image Generation

TILDE: TILt-based Distributional Erasure for Concept Unlearning

2026-07-07 · Naveen George, Naoki Murata, Yuhta Takida, Konda Reddy Mopuri 외 arxiv

Concept unlearning in text-to-image diffusion models is critical for safe and practical deployment: with rising privacy concerns, copyright disputes, trademark constraints, and safety regulations, deployed systems must b…

Beyond Text Prompts: Precise Concept Erasure through Text-Image Collaboration

2026-04-17 · Jun Li, Lizhi Xiong, Ziqiang Li, Weiwei Jiang 외 arxiv

Text-to-image generative models have achieved impressive fidelity and diversity, but can inadvertently produce unsafe or undesirable content due to implicit biases embedded in large-scale training datasets. Existing conc…

Text-to-Image GenerationRepresentation Learning