paper-with-me

홈 › Papers

PrivaCI-Bench: Evaluating Privacy with Contextual Integrity and Legal Compliance

2025-02-24 · Haoran Li, Wenbin Hu, Huihao Jing, Yulin Chen, Qi Hu, Sirui Han, Tianshu Chu, Peizhao Hu, Yangqiu Song

Recent advancements in generative large language models (LLMs) have enabled wider applicability, accessibility, and flexibility. However, their reliability and trustworthiness are still in doubt, especially for concerns regarding individuals' data privacy. Great efforts have been made on privacy by building various evaluation benchmarks to study LLMs' privacy awareness and robustness from their generated outputs to their hidden representations. Unfortunately, most of these works adopt a narrow formulation of privacy and only investigate personally identifiable information (PII). In this paper, we follow the merit of the Contextual Integrity (CI) theory, which posits that privacy evaluation should not only cover the transmitted attributes but also encompass the whole relevant social context through private information flows. We present PrivaCI-Bench, a comprehensive contextual privacy evaluation benchmark targeted at legal compliance to cover well-annotated privacy and safety regulations, real court cases, privacy policies, and synthetic data built from the official toolkit to study LLMs' privacy and safety compliance. We evaluate the latest LLMs, including the recent reasoner models QwQ-32B and Deepseek R1. Our experimental results suggest that though LLMs can effectively capture key CI parameters inside a given context, they still require further advancements for privacy compliance.

📄 PDF Abstract BibTeX arXiv:2502.17041

Code (1)

HKUST-KnowComp/PrivaCI-Bench 공식 구현 pytorch

Methods 이 논문이 사용한 방법론

ADOPT Please enter a description about the method here

Similar Papers 제목 키워드 기반

MPCI-Bench: A Benchmark for Multimodal Pairwise Contextual Integrity Evaluation of Language Model Agents

2026-01-13 · Shouju Wang, Haopeng Zhang arxiv

As language-model agents evolve from passive chatbots into proactive assistants that handle personal data, evaluating their adherence to social norms becomes increasingly critical, often through the lens of Contextual In…

Need to Know: Contextual-Integrity-Grounded Query Rewriting for Privacy-Conscious LLM Delegation

2026-06-02 · Xinyue Huang, Xiaochun Cao, Wenyuan Yang arxiv

As LLMs become increasingly woven into everyday workflows, user queries sent to cloud hosted LLMs routinely mix task-essential content with task non-essential sensitive disclosures, yet type based PII redaction is contex…

Reinforcement Learning

CI-Bench: Benchmarking Contextual Integrity of AI Assistants on Synthetic Data

2024-09-20 · Zhao Cheng, Diane Wan, Matthew Abueg, Sahra Ghalebikesabi 외

Advances in generative AI point towards a new era of personalized applications that perform diverse tasks on behalf of users. While general AI assistants have yet to fully emerge, their potential to share personal data r…

BenchmarkingLanguage ModelingLanguage Modelling

Do Vision-Language Models Respect Contextual Integrity in Location Disclosure?

2026-02-04 · Ruixin Yang, Ethan Mendes, Arthur Wang, James Hays 외 arxiv

Vision-language models (VLMs) have demonstrated strong performance in image geolocation, a capability further sharpened by frontier multimodal large reasoning models (MLRMs). This poses a significant privacy risk, as the…

AgentSCOPE: Evaluating Contextual Privacy Across Agentic Workflows

2026-03-05 · Ivoline C. Ngong, Keerthiram Murugesan, Swanand Kadhe, Justin D. Weisz 외 arxiv

Agentic systems are increasingly acting on users' behalf, accessing calendars, email, and personal files to complete everyday tasks. Privacy evaluation for these systems has focused on the input and output boundaries, bu…