paper-with-me

홈 › Papers

Roots Beneath the Cut: Uncovering the Risk of Concept Revival in Pruning-Based Unlearning for Diffusion Models

2026-02-28 · Ci Zhang, Zhaojun Ding, Chence Yang, Jun Liu, Xiaoming Zhai, Shaoyi Huang, Beiwen Li, Xiaolong Ma, Jin Lu, Geng Yuan arxiv

Pruning-based unlearning has recently emerged as a fast, training-free, and data-independent approach to remove undesired concepts from diffusion models. It promises high efficiency and robustness, offering an attractive alternative to traditional fine-tuning or editing-based unlearning. However, in this paper we uncover a hidden danger behind this promising paradigm. We find that the locations of pruned weights, typically set to zero during unlearning, can act as side-channel signals that leak critical information about the erased concepts. To verify this vulnerability, we design a novel attack framework capable of reviving erased concepts from pruned diffusion models in a fully data-free and training-free manner. Our experiments confirm that pruning-based unlearning is not inherently secure, as erased concepts can be effectively revived without any additional data or retraining. Extensive experiments on diffusion-based unlearning based on concept related weights lead to the conclusion: once the critical concept-related weights in diffusion models are identified, our method can effectively recover the original concept regardless of how the weights are manipulated. Finally, we explore potential defense strategies and advocate safer pruning mechanisms that conceal pruning locations while preserving unlearning effectiveness, providing practical insights for designing more secure pruning-based unlearning frameworks.

📄 PDF Abstract BibTeX arXiv:2603.06640

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Projected Gradient Unlearning for Text-to-Image Diffusion Models: Defending Against Concept Revival Attacks

2026-04-22 · Aljalila Aladawi, Mohammed Talha Alam, Fakhri Karray arxiv

Machine unlearning for text-to-image diffusion models aims to selectively remove undesirable concepts from pre-trained models without costly retraining. Current unlearning methods share a common weakness: erased concepts…

On the Principles of Parsimony and Self-Consistency for the Emergence of Intelligence

2022-07-11 · Yi Ma, Doris Tsao, Heung-Yeung Shum

Ten years into the revival of deep networks and artificial intelligence, we propose a theoretical framework that sheds light on understanding deep networks within a bigger picture of Intelligence in general. We introduce…

The role of causality in explainable artificial intelligence

2023-09-18 · Gianluca Carloni, Andrea Berti, Sara Colantonio

Causality and eXplainable Artificial Intelligence (XAI) have developed as separate fields in computer science, even though the underlying concepts of causation and explanation share common ancient roots. This is further …

Causal DiscoveryCausal IdentificationCausal InferenceExplainable artificial intelligence+4

Revival: Collaborative Artistic Creation through Human-AI Interactions in Musical Creativity

2025-01-19 · Keon Ju M. Lee, Philippe Pasquier, Jun Yuri

Revival is an innovative live audiovisual performance and music improvisation by our artist collective K-Phi-A, blending human and AI musicianship to create electronic music with audio-reactive visuals. The performance f…

A clarification of misconceptions, myths and desired status of artificial intelligence

2020-08-03 · Frank Emmert-Streib, Olli Yli-Harja, Matthias Dehmer

The field artificial intelligence (AI) has been founded over 65 years ago. Starting with great hopes and ambitious goals the field progressed though various stages of popularity and received recently a revival in the for…

BIG-bench Machine LearningMisconceptions