paper-with-me

홈 › Papers

Paradox of De-identification: A Critique of HIPAA Safe Harbour in the Age of LLMs

2026-02-09 · Lavender Y. Jiang, Xujin Chris Liu, Kyunghyun Cho, Eric K. Oermann arxiv

Privacy is a human right that sustains patient-provider trust. Clinical notes capture a patient's private vulnerability and individuality, which are used for care coordination and research. Under HIPAA Safe Harbor, these notes are de-identified to protect patient privacy. However, Safe Harbor was designed for an era of categorical tabular data, focusing on the removal of explicit identifiers while ignoring the latent information found in correlations between identity and quasi-identifiers, which can be captured by modern LLMs. We first formalize these correlations using a causal graph, then validate it empirically through individual re-identification of patients from scrubbed notes. The paradox of de-identification is further shown through a diagnosis ablation: even when all other information is removed, the model can predict the patient's neighborhood based on diagnosis alone. This position paper raises the question of how we can act as a community to uphold patient-provider trust when de-identification is inherently imperfect. We aim to raise awareness and discuss actionable recommendations.

📄 PDF Abstract BibTeX arXiv:2602.08997

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Deep classification algorithm for De-identification of DICOM medical images

2025-08-04 · Bufano Michele, Kotter Elmar arxiv

Background : De-identification of DICOM (Digital Imaging and Communi-cations in Medicine) files is an essential component of medical image research. Personal Identifiable Information (PII) and/or Personal Health Identify…

Fair Clustering: A Causal Perspective

2023-12-14 · Fritz Bayer, Drago Plecko, Niko Beerenwinkel, Jack Kuipers

Clustering algorithms may unintentionally propagate or intensify existing disparities, leading to unfair representations or biased decision-making. Current fair clustering methods rely on notions of fairness that do not …

ClusteringDecision MakingFairness

SAFETY-J: Evaluating Safety with Critique

2024-07-24 · Yixiu Liu, Yuxiang Zheng, Shijie Xia, Jiajun Li 외

The deployment of Large Language Models (LLMs) in content generation raises significant safety concerns, particularly regarding the transparency and interpretability of content evaluations. Current methods, primarily foc…

MetaSC: Test-Time Safety Specification Optimization for Language Models

2025-02-11 · Víctor Gallego

We propose a novel dynamic safety framework that optimizes language model (LM) safety reasoning at inference time without modifying model weights. Building on recent advances in self-critique methods, our approach levera…

Language ModelingLanguage Modelling

Agentic AI for Commercial Insurance Underwriting with Adversarial Self-Critique

2026-01-21 · Joyjit Roy, Samaresh Kumar Singh arxiv

Commercial insurance underwriting is a labor-intensive process that requires manual review of extensive documentation to assess risk and determine policy pricing. While AI offers substantial efficiency improvements, exis…