paper-with-me

홈 › Papers

MAGPIE: A benchmark for Multi-AGent contextual PrIvacy Evaluation

2025-10-16 · Gurusha Juneja, Jayanth Naga Sai Pasupulati, Alon Albalak, Wenyue Hua, William Yang Wang arxiv

A core challenge for autonomous LLM agents in collaborative settings is balancing robust privacy understanding and preservation alongside task efficacy. Existing privacy benchmarks only focus on simplistic, single-turn interactions where private information can be trivially omitted without affecting task outcomes. In this paper, we introduce MAGPIE (Multi-AGent contextual PrIvacy Evaluation), a novel benchmark of 200 high-stakes tasks designed to evaluate privacy understanding and preservation in multi-agent collaborative, non-adversarial scenarios. MAGPIE integrates private information as essential for task resolution, forcing agents to balance effective collaboration with strategic information control. Our evaluation reveals that state-of-the-art agents, including GPT-5 and Gemini 2.5-Pro, exhibit significant privacy leakage, with Gemini 2.5-Pro leaking up to 50.7% and GPT-5 up to 35.1% of the sensitive information even when explicitly instructed not to. Moreover, these agents struggle to achieve consensus or task completion and often resort to undesirable behaviors such as manipulation and power-seeking (e.g., Gemini 2.5-Pro demonstrating manipulation in 38.2% of the cases). These findings underscore that current LLM agents lack robust privacy understanding and are not yet adequately aligned to simultaneously preserve privacy and maintain effective collaboration in complex environments.

📄 PDF Abstract BibTeX arXiv:2510.15186

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

MAGPIE: A dataset for Multi-AGent contextual PrIvacy Evaluation

2025-06-25 · Gurusha Juneja, Alon Albalak, Wenyue Hua, William Yang Wang

The proliferation of LLM-based agents has led to increasing deployment of inter-agent collaboration for tasks like scheduling, negotiation, resource allocation etc. In such systems, privacy is critical, as agents often a…

Scheduling

PrivAct: Internalizing Contextual Privacy Preservation via Multi-Agent Preference Training

2026-02-14 · Yuhan Cheng, Hancheng Ye, Hai Helen Li, Jingwei Sun 외 arxiv

Large language model (LLM) agents are increasingly deployed in personalized tasks involving sensitive, context-dependent information, where privacy violations may arise in agents' action due to the implicitness of contex…

Zero-shot Generalization

MAGPIE: Multi-Task Media-Bias Analysis Generalization for Pre-Trained Identification of Expressions

2024-02-27 · Tomáš Horych, Martin Wessel, Jan Philip Wahle, Terry Ruas 외

Media bias detection poses a complex, multifaceted problem traditionally tackled using single-task models and small in-domain datasets, consequently lacking generalizability. To address this, we introduce MAGPIE, the fir…

Bias DetectionFake News Detection

LongMagpie: A Self-synthesis Method for Generating Large-scale Long-context Instructions

2025-05-22 · Chaochen Gao, Xing Wu, Zijia Lin, Debing Zhang 외

High-quality long-context instruction data is essential for aligning long-context large language models (LLMs). Despite the public release of models like Qwen and Llama, their long-context instruction data remains propri…

Diversity

1-2-3 Check: Enhancing Contextual Privacy in LLM via Multi-Agent Reasoning

2025-08-11 · Wenkai Li, Liwen Sun, Zhenxiang Guan, Xuhui Zhou 외 arxiv

Addressing contextual privacy concerns remains challenging in interactive settings where large language models (LLMs) process information from multiple sources (e.g., summarizing meetings with private and public informat…