paper-with-me

홈 › Papers

On Privacy and Confidentiality of Communications in Organizational Graphs

2021-05-27 · Masoumeh Shafieinejad, Huseyin Inan, Marcello Hasegawa, Robert Sim

Machine learned models trained on organizational communication data, such as emails in an enterprise, carry unique risks of breaching confidentiality, even if the model is intended only for internal use. This work shows how confidentiality is distinct from privacy in an enterprise context, and aims to formulate an approach to preserving confidentiality while leveraging principles from differential privacy. The goal is to perform machine learning tasks, such as learning a language model or performing topic analysis, using interpersonal communications in the organization, while not learning about confidential information shared in the organization. Works that apply differential privacy techniques to natural language processing tasks usually assume independently distributed data, and overlook potential correlation among the records. Ignoring this correlation results in a fictional promise of privacy. Naively extending differential privacy techniques to focus on group privacy instead of record-level privacy is a straightforward approach to mitigate this issue. This approach, although providing a more realistic privacy-guarantee, is over-cautious and severely impacts model utility. We show this gap between these two extreme measures of privacy over two language tasks, and introduce a middle-ground solution. We propose a model that captures the correlation in the social network graph, and incorporates this correlation in the privacy calculations through Pufferfish privacy principles.

📄 PDF Abstract BibTeX arXiv:2105.13418

Code (0)

등록된 구현이 없습니다.

Tasks

Language Modelling

Similar Papers 제목 키워드 기반

Privacy Meets Explainability: Managing Confidential Data and Transparency Policies in LLM-Empowered Science

2025-04-14 · Yashothara Shanmugarasa, Shidong Pan, Ming Ding, Dehai Zhao 외

As Large Language Models (LLMs) become integral to scientific workflows, concerns over the confidentiality and ethical handling of confidential data have emerged. This paper explores data exposure risks through LLM-power…

Federated Conformance Checking

2025-01-23 · Majid Rafiei, Mahsa Pourbafrani, Wil M. P. van der Aalst

Conformance checking is a crucial aspect of process mining, where the main objective is to compare the actual execution of a process, as recorded in an event log, with a reference process model, e.g., in the form of a Pe…

Chemical knowledge-informed framework for privacy-aware retrosynthesis learning

2025-02-26 · Guikun Chen, Xu Zhang, Xiaolin Hu, Yong liu 외

Chemical reaction data is a pivotal asset, driving advances in competitive fields such as pharmaceuticals, materials science, and industrial chemistry. Its proprietary nature renders it sensitive, as it often includes co…

Privacy PreservingRetrosynthesis

An Efficient Privacy-Preserving Multi-Keyword Query Scheme in Location Based Services

2020-08-21 · IEEE 2020 8 · SHIWEN ZHANG 1, 2, (Member, TINGTING YAO3 외

With the proliferation of location-aware mobile devices and the prevalence of wireless communications, location-based services (LBS) have attracted much particular attention in recent years. For flexibility and cost sa…

Privacy Preserving

Protecting Confidentiality, Privacy and Integrity in Collaborative Learning

2024-12-11 · Dong Chen, Alice Dethise, Istemi Ekin Akkus, Ivica Rimac 외

A collaboration between dataset owners and model owners is needed to facilitate effective machine learning (ML) training. During this collaboration, however, dataset owners and model owners want to protect the confidenti…

CPUGPUPrivacy Preserving