paper-with-me

홈 › Papers

Safeguarding Text-to-Image Generative Models Against Unauthorized Knowledge Distillation

2026-05-21 · Yilan Gao, Sida Huang, Hongyuan Zhang, Xuelong Li arxiv

Closed-weight generative services are increasingly deployed through query-based APIs, where users can obtain generated outputs while model parameters remain inaccessible. However, such deployment does not prevent model stealing: an attacker can repeatedly query the service, collect large volumes of released synthetic images, and use them as training data for a private substitute model. This query-output-driven process enables unauthorized knowledge distillation and capability replication without direct access to the original weights. To mitigate this threat, a practical defense should preserve the visual fidelity of released images, provide explicit control over perturbation magnitude, and scale efficiently to large-volume output release. We present WaveGuard, a single-pass, generator-based protection framework that safeguards released synthetic images under a user-specified perturbation budget. WaveGuard employs a frequency-aware perturbation generator to inject structured, imperceptible perturbations that maintain perceptual utility for benign viewers while reducing the usefulness of protected images as training data for unauthorized student models. Extensive experiments under WikiArt-related synthetic-output distillation settings show that WaveGuard achieves a favorable efficacy--fidelity--efficiency trade-off, with explicit imperceptibility control and substantial gains in protection efficiency.

📄 PDF Abstract BibTeX arXiv:2605.22060

Code (0)

등록된 구현이 없습니다.

Tasks

Knowledge Distillation

Similar Papers 제목 키워드 기반

Latent Diffusion Unlearning: Protecting Against Unauthorized Personalization Through Trajectory Shifted Perturbations

2025-10-03 · Naresh Kumar Devulapally, Shruti Agarwal, Tejas Gokhale, Vishnu Suresh Lokhande arxiv

Text-to-image diffusion models have demonstrated remarkable effectiveness in rapid and high-fidelity personalization, even when provided with only a few user images. However, the effectiveness of personalization techniqu…

Generative AI based Secure Wireless Sensing for ISAC Networks

2024-08-21 · Jiacheng Wang, Hongyang Du, Yinqiu Liu, Geng Sun 외

Integrated sensing and communications (ISAC) is expected to be a key technology for 6G, and channel state information (CSI) based sensing is a key component of ISAC. However, current research on ISAC focuses mainly on im…

Activity RecognitionISAC

Safeguarding Medical Image Segmentation Datasets against Unauthorized Training via Contour- and Texture-Aware Perturbations

2024-03-21 · Xun Lin, Yi Yu, Song Xia, Jue Jiang 외

The widespread availability of publicly accessible medical images has significantly propelled advancements in various research and clinical fields. Nonetheless, concerns regarding unauthorized training of AI systems for …

image-classificationImage ClassificationImage GenerationImage Segmentation+3

Do Not Merge My Model! Safeguarding Open-Source LLMs Against Unauthorized Model Merging

2025-11-13 · Qinfeng Li, Miao Pan, Jintao Chen, Fu Teng 외 arxiv

Model merging has emerged as an efficient technique for expanding large language models (LLMs) by integrating specialized expert models. However, it also introduces a new threat: model merging stealing, where free-riders…

Boosting Digital Safeguards: Blending Cryptography and Steganography

2024-04-09 · Anamitra Maiti, Subham Laha, Rishav Upadhaya, Soumyajit Biswas 외

In today's digital age, the internet is essential for communication and the sharing of information, creating a critical need for sophisticated data security measures to prevent unauthorized access and exploitation. Crypt…