paper-with-me

홈 › Papers

Untargeted Backdoor Watermark: Towards Harmless and Stealthy Dataset Copyright Protection

2022-09-27 · Yiming Li, Yang Bai, Yong Jiang, Yong Yang, Shu-Tao Xia, Bo Li

Deep neural networks (DNNs) have demonstrated their superiority in practice. Arguably, the rapid development of DNNs is largely benefited from high-quality (open-sourced) datasets, based on which researchers and developers can easily evaluate and improve their learning methods. Since the data collection is usually time-consuming or even expensive, how to protect their copyrights is of great significance and worth further exploration. In this paper, we revisit dataset ownership verification. We find that existing verification methods introduced new security risks in DNNs trained on the protected dataset, due to the targeted nature of poison-only backdoor watermarks. To alleviate this problem, in this work, we explore the untargeted backdoor watermarking scheme, where the abnormal model behaviors are not deterministic. Specifically, we introduce two dispersibilities and prove their correlation, based on which we design the untargeted backdoor watermark under both poisoned-label and clean-label settings. We also discuss how to use the proposed untargeted backdoor watermark for dataset ownership verification. Experiments on benchmark datasets verify the effectiveness of our methods and their resistance to existing backdoor defenses. Our codes are available at \url{https://github.com/THUYimingLi/Untargeted_Backdoor_Watermark}.

📄 PDF Abstract BibTeX arXiv:2210.00875

Code (1)

thuyimingli/untargeted_backdoor_watermark 공식 구현 pytorch

Similar Papers 제목 키워드 기반

StealthMark: Harmless and Stealthy Ownership Verification for Medical Segmentation via Uncertainty-Guided Backdoors

2026-01-23 · Qinkai Yu, Chong Zhang, Gaojie Jin, Tianjin Huang 외 arxiv

Annotating medical data for training AI models is often costly and limited due to the shortage of specialists with relevant clinical expertise. This challenge is further compounded by privacy and ethical concerns associa…

Towards Backdoor-Based Ownership Verification for Vision-Language-Action Models

2026-05-09 · Ming Sun, Rui Wang, Xingrui Yu, Lihua Jing 외 arxiv

Vision-Language-Action models (VLAs) support generalist robotic control by enabling end-to-end decision policies directly from multi-modal inputs. As trained VLAs are increasingly shared and adapted, protecting model own…

EntropyMark: Towards More Harmless Backdoor Watermark via Entropy-based Constraint for Open-source Dataset Copyright Protection

2025-01-01 · CVPR 2025 1 · Ming Sun, Rui Wang, Zixuan Zhu, Lihua Jing 외

High-quality open-source datasets are essential for advancing deep neural networks. However, the unauthorized commercial use of these datasets has raised significant concerns about copyright protection. One promising…

Prediction

Robust Watermarks Meet Backdoored Models: Evading Diffusion Semantic Watermarks via Stealthy Backdoor

2026-08-01 · Jinyuan Liu, Tianshuo Cong, Pei Li, Tianrui Wang 외 arxiv

Although semantic watermarking is considered a promising safeguard for images generated by Latent Diffusion Models (LDMs), the reliance of the watermark detection pipeline on neural networks introduces a critical yet und…

Data Taggants: Dataset Ownership Verification via Harmless Targeted Data Poisoning

2024-10-09 · Wassim Bouaziz, El-Mahdi El-Mhamdi, Nicolas Usunier

Dataset ownership verification, the process of determining if a dataset is used in a model's training data, is necessary for detecting unauthorized data usage and data contamination. Existing approaches, such as backdoor…

Data Poisoning