paper-with-me

홈 › Papers

Invisible Watermarks, Visible Gains: Steering Machine Unlearning with Bi-Level Watermarking Design

2025-08-13 · Yuhao Sun, Yihua Zhang, Gaowen Liu, Hongtao Xie, Sijia Liu arxiv

With the increasing demand for the right to be forgotten, machine unlearning (MU) has emerged as a vital tool for enhancing trust and regulatory compliance by enabling the removal of sensitive data influences from machine learning (ML) models. However, most MU algorithms primarily rely on in-training methods to adjust model weights, with limited exploration of the benefits that data-level adjustments could bring to the unlearning process. To address this gap, we propose a novel approach that leverages digital watermarking to facilitate MU by strategically modifying data content. By integrating watermarking, we establish a controlled unlearning mechanism that enables precise removal of specified data while maintaining model utility for unrelated tasks. We first examine the impact of watermarked data on MU, finding that MU effectively generalizes to watermarked data. Building on this, we introduce an unlearning-friendly watermarking framework, termed Water4MU, to enhance unlearning effectiveness. The core of Water4MU is a bi-level optimization (BLO) framework: at the upper level, the watermarking network is optimized to minimize unlearning difficulty, while at the lower level, the model itself is trained independently of watermarking. Experimental results demonstrate that Water4MU is effective in MU across both image classification and image generation tasks. Notably, it outperforms existing methods in challenging MU scenarios, known as "challenging forgets".

📄 PDF Abstract BibTeX arXiv:2508.10065

Code (0)

등록된 구현이 없습니다.

Tasks

Image ClassificationImage Generation

Similar Papers 제목 키워드 기반

Invisible Image Watermarks Are Provably Removable Using Generative AI

2023-06-02 · Xuandong Zhao, Kexun Zhang, Zihao Su, Saastha Vasan 외

Invisible watermarks safeguard images' copyrights by embedding hidden messages only detectable by owners. They also prevent people from misusing images, especially those generated by AI models. We propose a family of reg…

DenoisingImage Denoising

Assessing the Efficacy of Invisible Watermarks in AI-Generated Medical Images

2024-02-05 · Xiaodan Xing, Huiyu Zhou, Yingying Fang, Guang Yang

AI-generated medical images are gaining growing popularity due to their potential to address the data scarcity challenge in the real world. However, the issue of accurate identification of these synthetic images, particu…

A Baseline Method for Removing Invisible Image Watermarks using Deep Image Prior

2025-02-19 · Hengyue Liang, Taihui Li, Ju Sun

Image watermarks have been considered a promising technique to help detect AI-generated content, which can be used to protect copyright or prevent fake image abuse. In this work, we present a black-box method for removin…

BenchmarkingMisinformation

RAVEN: Erasing Invisible Watermarks via Novel View Synthesis

2026-01-13 · Fahad Shamshad, Nils Lukas, Karthik Nandakumar arxiv

Invisible watermarking has become a critical mechanism for authenticating AI-generated image content, with major platforms deploying watermarking schemes at scale. However, evaluating the vulnerability of these schemes a…

Novel View Synthesis

WMCopier: Forging Invisible Image Watermarks on Arbitrary Images

2025-03-28 · Ziping Dong, Chao Shuai, Zhongjie Ba, Peng Cheng 외

Invisible Image Watermarking is crucial for ensuring content provenance and accountability in generative AI. While Gen-AI providers are increasingly integrating invisible watermarking systems, the robustness of these sch…