paper-with-me

홈 › Papers

MuPPET: A Benchmark for Contextual Privacy of LLM Assistants in Multi-Party Conversations

2026-06-22 · Elena Sofia Ruzzetti, Cornelius Emde, Sangdoo Yun, Seong Joon Oh, Martin Gubri arxiv

LLM agents are increasingly deployed in multi-party environments, handling sensitive personal data on behalf of individual users, for instance in group chats. When such an agent discloses private information, it reaches every group member at once. This risk is structurally harder to control than in one-to-one settings, as every piece of private information must be appropriate for every recipient in the group. Yet all existing contextual privacy benchmarks consider only single-interlocutor settings, leaving multi-party privacy risks unmeasured. We introduce MuPPET (Multi-Party Privacy Exposure Testing), a benchmark for contextual privacy in multi-party conversations. Our experiments show that models leak substantially more in multi-party settings than one-to-one evaluations suggest. Frontier models are vulnerable, and smaller open-weights models, often preferred for local deployment with sensitive data, even more so. Existing contextual privacy defences offer only partial protection, degrade utility, and do not resolve the underlying party-tracking problem.

📄 PDF Abstract BibTeX arXiv:2606.23217

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Operationalizing Contextual Integrity in Privacy-Conscious Assistants

2024-08-05 · Sahra Ghalebikesabi, Eugene Bagdasaryan, Ren Yi, Itay Yona 외

Advanced AI assistants combine frontier LLMs and tool access to autonomously perform complex tasks on behalf of users. While the helpfulness of such assistants can increase dramatically with access to user information in…

CI-Bench: Benchmarking Contextual Integrity of AI Assistants on Synthetic Data

2024-09-20 · Zhao Cheng, Diane Wan, Matthew Abueg, Sahra Ghalebikesabi 외

Advances in generative AI point towards a new era of personalized applications that perform diverse tasks on behalf of users. While general AI assistants have yet to fully emerge, their potential to share personal data r…

BenchmarkingLanguage ModelingLanguage Modelling

MPCI-Bench: A Benchmark for Multimodal Pairwise Contextual Integrity Evaluation of Language Model Agents

2026-01-13 · Shouju Wang, Haopeng Zhang arxiv

As language-model agents evolve from passive chatbots into proactive assistants that handle personal data, evaluating their adherence to social norms becomes increasingly critical, often through the lens of Contextual In…

3D-MuPPET: 3D Multi-Pigeon Pose Estimation and Tracking

2023-08-29 · Urs Waldmann, Alex Hoi Hang Chan, Hemal Naik, Máté Nagy 외

Markerless methods for animal posture tracking have been rapidly developing recently, but frameworks and benchmarks for tracking large animal groups in 3D are still lacking. To overcome this gap in the literature, we pre…

Pose Estimation

Can LLMs Keep a Secret? Testing Privacy Implications of Language Models via Contextual Integrity Theory

2023-10-27 · Niloofar Mireshghallah, Hyunwoo Kim, Xuhui Zhou, Yulia Tsvetkov 외

The interactive use of large language models (LLMs) in AI assistants (at work, home, etc.) introduces a new set of inference-time privacy risks: LLMs are fed different types of information from multiple sources in their …

Privacy Preserving