paper-with-me

Papers

What Do LLMs Associate with Your Name? A Human-Centered Black-Box Audit of Personal Data

2026-02-19 · Dimitri Staufer, Kirsten Morehouse arxiv

Large language models (LLMs), and conversational agents based on them, are exposed to personal data (PD) during pre-training and during user interactions. Prior work shows that PD can resurface, yet users lack insight into how strongly models associate specific information to their identity. We audit PD across eight LLMs (3 open-source; 5 API-based, including GPT-4o), introduce LMP2 (Language Model Privacy Probe), a human-centered, privacy-preserving audit tool refined through two formative studies (N=20), and run two studies with EU residents to capture (i) intuitions about LLM-generated PD (N1=155) and (ii) reactions to tool output (N2=303). We show empirically that models confidently generate multiple PD categories for well-known individuals. For everyday users, GPT-4o generates 11 features with 60% or more accuracy (e.g., gender, hair color, languages). Finally, 72% of participants sought control over model-generated associations with their name, raising questions about what counts as PD and whether data privacy rights should extend to LLMs.

📄 PDF Abstract BibTeX arXiv:2602.17483

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Investigating Agency of LLMs in Human-AI Collaboration Tasks

2023-05-22 · ASHISH SHARMA, Sudha Rao, Chris Brockett, Akanksha Malhotra 외

Agency, the capacity to proactively shape events, is central to how humans interact and collaborate. While LLMs are being developed to simulate human behavior and serve as human-like agents, little attention has been giv…

Can Large Language Models Really Recognize Your Name?

2025-05-20 · Dzung Pham, Peter Kairouz, Niloofar Mireshghallah, Eugene Bagdasarian 외

Large language models (LLMs) are increasingly being used to protect sensitive user data. However, current LLM-based privacy solutions assume that these models can reliably detect personally identifiable information (PII)…

Privacy Preserving

Human-Centred LLM Privacy Audits: Findings and Frictions

2026-03-12 · Dimitri Staufer, Kirsten Morehouse, David Hartmann, Bettina Berendt arxiv

Large language models (LLMs) learn statistical associations from massive training corpora and user interactions, and deployed systems can surface or infer information about individuals. Yet people lack practical ways to …

What Your Username Says About You

2015-07-08 · EMNLP 2015 9 · Aaron Jaech, Mari Ostendorf

Usernames are ubiquitous on the Internet, and they are often suggestive of user demographics. This work looks at the degree to which gender and language can be inferred from a username alone by making use of unsupervised…

What makes your model a low-empathy or warmth person: Exploring the Origins of Personality in LLMs

2024-10-07 · Shu Yang, Shenzhe Zhu, Ruoxuan Bao, Liang Liu 외

Large language models (LLMs) have demonstrated remarkable capabilities in generating human-like text and exhibiting personality traits similar to those in humans. However, the mechanisms by which LLMs encode and express …