paper-with-me

홈 › Papers

Erasure or Erosion? Evaluating Compositional Degradation in Unlearned Text-To-Image Diffusion Models

2026-04-06 · Arian Komaei Koma, Seyed Amir Kasaei, Ali Aghayari, AmirMahdi Sadeghzadeh, Mohammad Hossein Rohban arxiv

Post-hoc unlearning has emerged as a practical mechanism for removing undesirable concepts from large text-to-image diffusion models. However, prior work primarily evaluates unlearning through erasure success; its impact on broader generative capabilities remains poorly understood. In this work, we conduct a systematic empirical study of concept unlearning through the lens of compositional text-to-image generation. Focusing on nudity removal in Stable Diffusion 1.4, we evaluate a diverse set of state-of-the-art unlearning methods using T2I-CompBench++ and GenEval, alongside established unlearning benchmarks. Our results reveal a consistent trade-off between unlearning effectiveness and compositional integrity: methods that achieve strong erasure frequently incur substantial degradation in attribute binding, spatial reasoning, and counting. Conversely, approaches that preserve compositional structure often fail to provide robust erasure. These findings highlight limitations of current evaluation practices and underscore the need for unlearning objectives that explicitly account for semantic preservation beyond targeted suppression.

📄 PDF Abstract BibTeX arXiv:2604.04575

Code (0)

등록된 구현이 없습니다.

Tasks

Text-to-Image GenerationSpatial Reasoning

Similar Papers 제목 키워드 기반

Mosaic: Compositional Multi-Concept Erasure via Vector Field Blending

2026-05-25 · Junseok Ko, Jungwoo Kim, Jong-Seok Lee arxiv

Concept erasure has emerged as a key research direction for ensuring safe and ethical image synthesis in Text-to-Image (T2I) models. While existing studies have explored concept erasure across multiple concepts, they typ…

On the Vulnerability of Concept Erasure in Diffusion Models

2025-02-24 · Lucas Beerens, Alex D. Richardson, Kaicheng Zhang, Dongdong Chen

The proliferation of text-to-image diffusion models has raised significant privacy and security concerns, particularly regarding the generation of copyrighted or harmful images. In response, several concept erasure (defe…

Machine Unlearning

SlopCodeBench: Benchmarking How Coding Agents Degrade Over Long-Horizon Iterative Tasks

2026-03-25 · Gabriel Orlanski, Devjeet Roy, Alexander Yun, Changho Shin 외 arxiv

Software development is iterative, yet agentic coding benchmarks hide design issues through their single-shot setup. Recent iterative benchmarks attempt to remedy this but heavily constrain an agent's design decision spa…

Beyond Fixed Anchors: Precisely Erasing Concepts with Sibling Exclusive Counterparts

2025-10-18 · Tong Zhang, Ru Zhang, Jianyi Liu, Zhen Yang 외 arxiv

Existing concept erasure methods for text-to-image diffusion models commonly rely on fixed anchor strategies, which often lead to critical issues such as concept re-emergence and erosion. To address this, we conduct caus…

Reliable and Efficient Concept Erasure of Text-to-Image Diffusion Models

2024-07-17 · Chao Gong, Kai Chen, Zhipeng Wei, Jingjing Chen 외

Text-to-image models encounter safety issues, including concerns related to copyright and Not-Safe-For-Work (NSFW) content. Despite several methods have been proposed for erasing inappropriate concepts from diffusion mod…

BenchmarkingRed Teaming