paper-with-me

Papers

CIDER: A Dataset of Contextual Disclosure Boundaries for Privacy Preference Alignment

2026-08-10 · Bingcan Guo, Eryue Xu, Jijie Zhou, Zhiping Zhang, Tianshi Li arxiv

Aligning large language models (LLMs) with human privacy preferences requires capturing individuals' disclosure boundaries beyond general privacy norms. However, a gap remains in eliciting such nuanced preferences to evaluate alignment in realistic settings. We introduce CIDER, a dataset of 14,850 human annotations from 169 users, forming 1,650 contextual disclosure boundary sets across 60 interpersonal communication scenarios involving information sharing that violates privacy norms. Each boundary represents a real user's disclosure decisions over 9 sharing variants in a scenario, for a given communication role and AI-mediated condition. We formulate a task in which models predict a user's disclosure decision from historical boundaries, with varying levels of contextual information. Across 12 open and proprietary models, in-context personalization improves prediction accuracy by up to 11.41 percentage points using only 6 historical examples. Larger models such as GPT-5.4 (with medium reasoning effort) and Claude Sonnet 4.6 are better at leveraging semantic context to understand user-specific, context-dependent disclosure preferences for more accurate predictions, while smaller models tend to rely on structured heuristics based on disclosure granularity and identifiability. Personalization generally improves prediction accuracy, but the improvement is often accompanied by imbalanced shifts in false-positive and false-negative rates across models, with only Claude Sonnet 4.6 achieving balanced improvements in both. Our findings reveal both the promise and limitations of inference-time personalization for privacy preference modeling and position CIDER as a resource for advancing personalized privacy alignment.

📄 PDF Abstract BibTeX arXiv:2608.09164

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Personalized Privacy Control in LLMs via Attention Head Intervention

2026-08-21 · Junseok Kim, Nakyeong Yang, Kyomin Jung arxiv

The rise of agentic AI enables LLMs to access diverse user data, raising critical privacy concerns. Prior work on contextual privacy studies whether LLMs regulate information disclosure according to context-dependent nor…

Not My Agent, Not My Boundary? Elicitation of Personal Privacy Boundaries in AI-Delegated Information Sharing

2025-09-26 · Bingcan Guo, Eryue Xu, Zhiping Zhang, Tianshi Li arxiv

Aligning AI systems with human privacy preferences requires understanding individuals' nuanced disclosure behaviors beyond general norms. Yet eliciting such boundaries remains challenging due to the context-dependent nat…

Do Vision-Language Models Respect Contextual Integrity in Location Disclosure?

2026-02-04 · Ruixin Yang, Ethan Mendes, Arthur Wang, James Hays 외 arxiv

Vision-language models (VLMs) have demonstrated strong performance in image geolocation, a capability further sharpened by frontier multimodal large reasoning models (MLRMs). This poses a significant privacy risk, as the…

Need to Know: Contextual-Integrity-Grounded Query Rewriting for Privacy-Conscious LLM Delegation

2026-06-02 · Xinyue Huang, Xiaochun Cao, Wenyuan Yang arxiv

As LLMs become increasingly woven into everyday workflows, user queries sent to cloud hosted LLMs routinely mix task-essential content with task non-essential sensitive disclosures, yet type based PII redaction is contex…

Reinforcement Learning

Privacy in Human-AI Romantic Relationships: Concerns, Boundaries, and Agency

2026-01-23 · Rongjun Ma, Shijing He, Jose Luis Martin-Navarro, Xiao Zhan 외 arxiv

An increasing number of LLM-based applications are being developed to facilitate romantic relationships with AI partners, yet the safety and privacy risks in these partnerships remain largely underexplored. In this work,…