paper-with-me

홈 › Papers

EntropyMark: Towards More Harmless Backdoor Watermark via Entropy-based Constraint for Open-source Dataset Copyright Protection

2025-01-01 · CVPR 2025 1 · Ming Sun, Rui Wang, Zixuan Zhu, Lihua Jing, Yuanfang Guo

High-quality open-source datasets are essential for advancing deep neural networks. However, the unauthorized commercial use of these datasets has raised significant concerns about copyright protection. One promising approach is backdoor watermark-based dataset ownership verification (BW-DOV), in which dataset protectors implant specific backdoors into illicit models through dataset watermarking, enabling the tracing of these models through abnormal prediction behaviors. Unfortunately, the targeted nature of these BW-DOV methods can be maliciously exploited, potentially leading to harmful side effects. While existing harmless methods attempt to mitigate these risks, watermarked datasets can still negatively affect prediction results, partially compromising dataset functionality. In this paper, we propose a more harmless backdoor watermark, called EntropyMark, which improves prediction confidence without altering the final prediction results. For this purpose, an entropy-based constraint is introduced to regulate the probability distribution. Specifically, we design an iterative clean-label dataset watermarking framework. Our framework employs gradient matching and adaptive data selection to optimize backdoor injection. In parallel, we introduce a hypothesis test method grounded in entropy inconsistency to verify dataset ownership. Extensive experiments on benchmark datasets demonstrate the effectiveness, transferability, and defense resistance of our approach.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Prediction

Similar Papers 제목 키워드 기반

Untargeted Backdoor Watermark: Towards Harmless and Stealthy Dataset Copyright Protection

2022-09-27 · Yiming Li, Yang Bai, Yong Jiang, Yong Yang 외

Deep neural networks (DNNs) have demonstrated their superiority in practice. Arguably, the rapid development of DNNs is largely benefited from high-quality (open-sourced) datasets, based on which researchers and develope…

Explanation as a Watermark: Towards Harmless and Multi-bit Model Ownership Verification via Watermarking Feature Attribution

2024-05-08 · Shuo Shao, Yiming Li, Hongwei Yao, Yiling He 외

Ownership verification is currently the most critical and widely adopted post-hoc method to safeguard model copyright. In general, model owners exploit it to identify whether a given suspicious third-party model is stole…

Explainable artificial intelligenceimage-classificationImage ClassificationText Generation

Towards Backdoor-Based Ownership Verification for Vision-Language-Action Models

2026-05-09 · Ming Sun, Rui Wang, Xingrui Yu, Lihua Jing 외 arxiv

Vision-Language-Action models (VLAs) support generalist robotic control by enabling end-to-end decision policies directly from multi-modal inputs. As trained VLAs are increasingly shared and adapted, protecting model own…

Double-I Watermark: Protecting Model Copyright for LLM Fine-tuning

2024-02-22 · Shen Li, Liuyi Yao, Jinyang Gao, Lan Zhang 외

To support various applications, a prevalent and efficient approach for business owners is leveraging their valuable datasets to fine-tune a pre-trained LLM through the API provided by LLM owners or cloud servers. Howeve…

StealthMark: Harmless and Stealthy Ownership Verification for Medical Segmentation via Uncertainty-Guided Backdoors

2026-01-23 · Qinkai Yu, Chong Zhang, Gaojie Jin, Tianjin Huang 외 arxiv

Annotating medical data for training AI models is often costly and limited due to the shortage of specialists with relevant clinical expertise. This challenge is further compounded by privacy and ethical concerns associa…