paper-with-me

Papers

Robust Membership Encoding: Inference Attacks and Copyright Protection for Deep Learning

2019-09-27 · Congzheng Song, Reza Shokri

Machine learning as a service (MLaaS), and algorithm marketplaces are on a rise. Data holders can easily train complex models on their data using third party provided learning codes. Training accurate ML models requires massive labeled data and advanced learning algorithms. The resulting models are considered as intellectual property of the model owners and their copyright should be protected. Also, MLaaS needs to be trusted not to embed secret information about the training data into the model, such that it could be later retrieved when the model is deployed. In this paper, we present \emph{membership encoding} for training deep neural networks and encoding the membership information, i.e. whether a data point is used for training, for a subset of training data. Membership encoding has several applications in different scenarios, including robust watermarking for model copyright protection, and also the risk analysis of stealthy data embedding privacy attacks. Our encoding algorithm can determine the membership of significantly redacted data points, and is also robust to model compression and fine-tuning. It also enables encoding a significant fraction of the training set, with negligible drop in the model's prediction accuracy.

📄 PDF Abstract BibTeX arXiv:1909.12982

Code (0)

등록된 구현이 없습니다.

Tasks

Deep LearningModel Compression

Similar Papers 제목 키워드 기반

Membership and Dataset Inference Attacks on Large Audio Generative Models

2025-12-10 · Jakub Proboszcz, Paweł Kochanski, Karol Korszun, Donato Crisostomi 외 arxiv

Generative audio models, based on diffusion and autoregressive architectures, have advanced rapidly in both quality and expressiveness. This progress, however, raises pressing copyright concerns, as such models are often…

CDI: Copyrighted Data Identification in Diffusion Models

2024-11-19 · CVPR 2025 1 · Jan Dubiński, Antoni Kowalczuk, Franziska Boenisch, Adam Dziedzic

Diffusion Models (DMs) benefit from large and diverse datasets for their training. Since this data is often scraped from the Internet without permission from the data owners, this raises concerns about copyright and inte…

Towards More Realistic Membership Inference Attacks on Large Diffusion Models

2023-06-22 · Jan Dubiński, Antoni Kowalczuk, Stanisław Pawlak, Przemysław Rokita 외

Generative diffusion models, including Stable Diffusion and Midjourney, can generate visually appealing, diverse, and high-resolution images for various applications. These models are trained on billions of internet-sour…

Inference AttackMembership Inference Attack

CopyrightMeter: Revisiting Copyright Protection in Text-to-image Models

2024-11-20 · Naen Xu, Changjiang Li, Tianyu Du, Minxi Li 외

Text-to-image diffusion models have emerged as powerful tools for generating high-quality images from textual descriptions. However, their increasing popularity has raised significant copyright concerns, as these models …

Image GenerationText to Image GenerationText-to-Image Generation

The Surprising Effectiveness of Membership Inference with Simple N-Gram Coverage

2025-08-13 · Skyler Hallinan, Jaehun Jung, Melanie Sclar, Ximing Lu 외 arxiv

Membership inference attacks serves as useful tool for fair use of language models, such as detecting potential copyright infringement and auditing data leakage. However, many current state-of-the-art attacks require acc…