paper-with-me

Papers

Can Copyright be Reduced to Privacy?

2023-05-24 · Niva Elkin-Koren, Uri Hacohen, Roi Livni, Shay Moran

There is a growing concern that generative AI models will generate outputs closely resembling the copyrighted materials for which they are trained. This worry has intensified as the quality and complexity of generative models have immensely improved, and the availability of extensive datasets containing copyrighted material has expanded. Researchers are actively exploring strategies to mitigate the risk of generating infringing samples, with a recent line of work suggesting to employ techniques such as differential privacy and other forms of algorithmic stability to provide guarantees on the lack of infringing copying. In this work, we examine whether such algorithmic stability techniques are suitable to ensure the responsible use of generative models without inadvertently violating copyright laws. We argue that while these techniques aim to verify the presence of identifiable information in datasets, thus being privacy-oriented, copyright law aims to promote the use of original works for the benefit of society as a whole, provided that no unlicensed use of protected expression occurred. These fundamental differences between privacy and copyright must not be overlooked. In particular, we demonstrate that while algorithmic stability may be perceived as a practical tool to detect copying, such copying does not necessarily constitute copyright infringement. Therefore, if adopted as a standard for detecting an establishing copyright infringement, algorithmic stability may undermine the intended objectives of copyright law.

📄 PDF Abstract BibTeX arXiv:2305.14822

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Privacy and Copyright Protection in Generative AI: A Lifecycle Perspective

2023-11-30 · Dawen Zhang, Boming Xia, Yue Liu, Xiwei Xu 외

The advent of Generative AI has marked a significant milestone in artificial intelligence, demonstrating remarkable capabilities in generating realistic images, texts, and data patterns. However, these advancements come …

Data PoisoningMachine Unlearning

Randomization Techniques to Mitigate the Risk of Copyright Infringement

2024-08-21 · Wei-Ning Chen, Peter Kairouz, Sewoong Oh, Zheng Xu

In this paper, we investigate potential randomization approaches that can complement current practices of input-based methods (such as licensing data and prompt filtering) and output-based methods (such as recitation che…

Blameless Users in a Clean Room: Defining Copyright Protection for Generative Models

2025-06-23 · Aloni Cohen

Are there any conditions under which a generative model's outputs are guaranteed not to infringe the copyrights of its training data? This is the question of "provable copyright protection" first posed by Vyas, Kakade, a…

counterfactual

Copyright Infringement Detection in Text-to-Image Diffusion Models via Differential Privacy

2025-09-27 · Xiafeng Man, Zhipeng Wei, Jingjing Chen arxiv

The widespread deployment of large vision models such as Stable Diffusion raises significant legal and ethical concerns, as these models can memorize and reproduce copyrighted content without authorization. Existing dete…

ISACL: Internal State Analyzer for Copyrighted Training Data Leakage

2025-08-25 · Guangwei Zhang, Qisheng Su, Jiateng Liu, Cheng Qian 외 arxiv

Large Language Models (LLMs) have revolutionized Natural Language Processing (NLP) but pose risks of inadvertently exposing copyrighted or proprietary data, especially when such data is used for training but not intended…

Text Generation