paper-with-me

Papers

Who Stole Your Data? A Method for Detecting Unauthorized RAG Theft

2025-10-09 · Peiyang Liu, Ziqiang Cui, Di Liang, Wei Ye arxiv

Retrieval-augmented generation (RAG) enhances Large Language Models (LLMs) by mitigating hallucinations and outdated information issues, yet simultaneously facilitates unauthorized data appropriation at scale. This paper addresses this challenge through two key contributions. First, we introduce RPD, a novel dataset specifically designed for RAG plagiarism detection that encompasses diverse professional domains and writing styles, overcoming limitations in existing resources. Second, we develop a dual-layered watermarking system that embeds protection at both semantic and lexical levels, complemented by an interrogator-detective framework that employs statistical hypothesis testing on accumulated evidence. Extensive experimentation demonstrates our approach's effectiveness across varying query volumes, defense prompts, and retrieval parameters, while maintaining resilience against adversarial evasion techniques. This work establishes a foundational framework for intellectual property protection in retrieval-augmented AI systems.

📄 PDF Abstract BibTeX arXiv:2510.07728

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Non-Intrusive Machine Learning Solution for Malware Detection and Data Theft Classification in Smartphones

2021-02-12 · Sai Vishwanath Venkatesh, Prasanna D. Kumaran, Joish J Bosco, Pravin R. Kumaar 외

Smartphones contain information that is more sensitive and personal than those found on computers and laptops. With an increase in the versatility of smartphone functionality, more data has become vulnerable and exposed …

BIG-bench Machine LearningMalware Detection

Watermarking One for All: A Robust Watermarking Scheme Against Partial Image Theft

2025-01-01 · CVPR 2025 1 · Gaozhi Liu, Silu Cao, Zhenxing Qian, Xinpeng Zhang 외

The proliferation of digital images on the Internet has provided unprecedented convenience, but also poses significant risks of malicious theft and misuse. Digital watermarking has long been researched as an effectiv…

All

Steal My Artworks for Fine-tuning? A Watermarking Framework for Detecting Art Theft Mimicry in Text-to-Image Models

2023-11-22 · Ge Luo, Junqiang Huang, Manman Zhang, Zhenxing Qian 외

The advancement in text-to-image models has led to astonishing artistic performances. However, several studios and websites illegally fine-tune these models using artists' artworks to mimic their styles for profit, which…

Garbage in, model out: Weight theft with just noise

2019-05-28 · Vinay Uday Prabhu, Nick Roberts, Matthew McAteer

This paper explores the scenarios under which an attacker can claim that ‘Noise and access to the softmax layer of the model is all you need’ to steal the weights of a convolutional neural network whose architecture is a…

Model Weight Theft With Just Noise Inputs: The Curious Case of the Petulant Attacker

2019-12-19 · Nicholas Roberts, Vinay Uday Prabhu, Matthew McAteer

This paper explores the scenarios under which an attacker can claim that 'Noise and access to the softmax layer of the model is all you need' to steal the weights of a convolutional neural network whose architecture is a…