paper-with-me

홈 › Papers

Towards Task-Agnostic Privacy- and Utility-Preserving Models

2021-09-01 · RANLP 2021 9 · Yaroslav Emelyanov

Modern deep learning models for natural language processing rely heavily on large amounts of annotated texts. However, obtaining such texts may be difficult when they contain personal or confidential information, for example, in health or legal domains. In this work, we propose a method of de-identifying free-form text documents by carefully redacting sensitive data in them. We show that our method preserves data utility for text classification, sequence labeling and question answering tasks.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Question Answeringtext-classificationText Classification

Similar Papers 제목 키워드 기반

SharedRequest: Privacy-Preserving Model-Agnostic Inference for Large Language Models

2026-06-03 · Peihua Mai, Xuanrong Gao, Youlong Ding, Xianglong Du 외 arxiv

With the widespread deployment of public large language models (LLMs) such as ChatGPT, protecting user prompt privacy has become an increasingly critical issue. Existing privacy-preserving inference methods sacrifice eit…

How to Train Private Clinical Language Models: A Comparative Study of Privacy-Preserving Pipelines for ICD-9 Coding

2025-11-18 · Mathieu Dufour, Andrew Duncan arxiv

Large language models trained on clinical text risk exposing sensitive patient information, yet differential privacy (DP) methods often severely degrade the diagnostic accuracy needed for deployment. Despite rapid progre…

Knowledge DistillationText Generation

Adaptive Clipping for Privacy-Preserving Few-Shot Learning: Enhancing Generalization with Limited Data

2025-03-27 · Kanishka Ranaweera, Dinh C. Nguyen, Pubudu N. Pathirana, David Smith 외

In the era of data-driven machine-learning applications, privacy concerns and the scarcity of labeled data have become paramount challenges. These challenges are particularly pronounced in the domain of few-shot learning…

Few-Shot LearningMeta-LearningPrivacy Preserving

Enhancing the Utility of Privacy-Preserving Cancer Classification using Synthetic Data

2024-07-17 · Richard Osuala, Daniel M. Lang, Anneliese Riess, Georgios Kaissis 외

Deep learning holds immense promise for aiding radiologists in breast cancer detection. However, achieving optimal model performance is hampered by limitations in availability and sharing of data commonly associated to p…

Breast Cancer DetectionCancer ClassificationData AugmentationDeep Learning+4

A Utility-preserving De-identification Pipeline for Cross-hospital Radiology Data Sharing

2026-04-08 · Chenhao Liu, Zelin Wen, Yan Tong, Junjie Zhu 외 arxiv

Large-scale radiology data are critical for developing robust medical AI systems. However, sharing such data across hospitals remains heavily constrained by privacy concerns. Existing de-identification research in radiol…