paper-with-me

Papers

Position: Privacy Is Not Just Memorization!

2025-10-02 · Niloofar Mireshghallah, Tianshi Li arxiv

The discourse on privacy risks in Large Language Models (LLMs) has disproportionately focused on verbatim memorization of training data, while a constellation of more immediate and scalable privacy threats remain underexplored. This position paper argues that the privacy landscape of LLM systems extends far beyond training data extraction, encompassing risks from data collection practices, inference-time context leakage, autonomous agent capabilities, and the democratization of surveillance through deep inference attacks. We present a comprehensive taxonomy of privacy risks across the LLM lifecycle -- from data collection through deployment -- and demonstrate through case studies how current privacy frameworks fail to address these multifaceted threats. Through a longitudinal analysis of 1,322 AI/ML privacy papers published at leading conferences over the past decade (2016--2025), we reveal that while memorization receives outsized attention in technical research, the most pressing privacy harms lie elsewhere, where current technical approaches offer little traction and viable paths forward remain unclear. We call for a fundamental shift in how the research community approaches LLM privacy, moving beyond the narrow focus of current technical solutions and embracing interdisciplinary approaches that address the sociotechnical nature of these emerging threats.

📄 PDF Abstract BibTeX arXiv:2510.01645

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

You Don't Need All That Attention: Surgical Memorization Mitigation in Text-to-Image Diffusion Models

2026-02-23 · Kairan Zhao, Eleni Triantafillou, Peter Triantafillou arxiv

Generative models have been shown to "memorize" certain training data, leading to verbatim or near-verbatim generating images, which may cause privacy concerns or copyright infringement. We introduce Guidance Using Attra…

Image GenerationImage Denoising

Rethinking Memorization Measures and their Implications in Large Language Models

2025-07-20 · Bishwamittra Ghosh, Soumi Das, Qinyuan Wu, Mohammad Aflah Khan 외 arxiv

Concerned with privacy threats, memorization in LLMs is often seen as undesirable, specifically for learning. In this paper, we study whether memorization can be avoided when optimally learning a language, and whether th…

Memorization in deep learning: A survey

2024-06-06 · Jiaheng Wei, Yanjun Zhang, Leo Yu Zhang, Ming Ding 외

Deep Learning (DL) powered by Deep Neural Networks (DNNs) has revolutionized various domains, yet understanding the intricacies of DNN decision-making and learning processes remains a significant challenge. Recent invest…

Decision MakingDeep LearningMemorizationSurvey

Unveiling Privacy, Memorization, and Input Curvature Links

2024-02-28 · Deepak Ravikumar, Efstathia Soufleri, Abolfazl Hashemi, Kaushik Roy

Deep Neural Nets (DNNs) have become a pervasive tool for solving many emerging problems. However, they tend to overfit to and memorize the training set. Memorization is of keen interest since it is closely related to sev…

Memorization

Towards Differential Relational Privacy and its use in Question Answering

2022-03-30 · Simone Bombari, Alessandro Achille, Zijian Wang, Yu-Xiang Wang 외

Memorization of the relation between entities in a dataset can lead to privacy issues when using a trained model for question answering. We introduce Relational Memorization (RM) to understand, quantify and control this …

MemorizationQuestion Answering