paper-with-me

홈 › Papers

Improving the Performance of Radiology Report De-identification with Large-Scale Training and Benchmarking Against Cloud Vendor Methods

2025-11-06 · Eva Prakash, Maayane Attias, Pierre Chambon, Justin Xu, Steven Truong, Jean-Benoit Delbrouck, Tessa Cook, Curtis Langlotz arxiv

Objective: To enhance automated de-identification of radiology reports by scaling transformer-based models through extensive training datasets and benchmarking performance against commercial cloud vendor systems for protected health information (PHI) detection. Materials and Methods: In this retrospective study, we built upon a state-of-the-art, transformer-based, PHI de-identification pipeline by fine-tuning on two large annotated radiology corpora from Stanford University, encompassing chest X-ray, chest CT, abdomen/pelvis CT, and brain MR reports and introducing an additional PHI category (AGE) into the architecture. Model performance was evaluated on test sets from Stanford and the University of Pennsylvania (Penn) for token-level PHI detection. We further assessed (1) the stability of synthetic PHI generation using a "hide-in-plain-sight" method and (2) performance against commercial systems. Precision, recall, and F1 scores were computed across all PHI categories. Results: Our model achieved overall F1 scores of 0.973 on the Penn dataset and 0.996 on the Stanford dataset, outperforming or maintaining the previous state-of-the-art model performance. Synthetic PHI evaluation showed consistent detectability (overall F1: 0.959 [0.958-0.960]) across 50 independently de-identified Penn datasets. Our model outperformed all vendor systems on synthetic Penn reports (overall F1: 0.960 vs. 0.632-0.754). Discussion: Large-scale, multimodal training improved cross-institutional generalization and robustness. Synthetic PHI generation preserved data utility while ensuring privacy. Conclusion: A transformer-based de-identification model trained on diverse radiology datasets outperforms prior academic and commercial systems in PHI detection and establishes a new benchmark for secure clinical text processing.

📄 PDF Abstract BibTeX arXiv:2511.04079

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Utility-preserving De-identification Pipeline for Cross-hospital Radiology Data Sharing

2026-04-08 · Chenhao Liu, Zelin Wen, Yan Tong, Junjie Zhu 외 arxiv

Large-scale radiology data are critical for developing robust medical AI systems. However, sharing such data across hospitals remains heavily constrained by privacy concerns. Existing de-identification research in radiol…

CheXpert Plus: Augmenting a Large Chest X-ray Dataset with Text Radiology Reports, Patient Demographics and Additional Image Formats

2024-05-29 · Pierre Chambon, Jean-Benoit Delbrouck, Thomas Sounack, Shih-Cheng Huang 외

Since the release of the original CheXpert paper five years ago, CheXpert has become one of the most widely used and cited clinical AI datasets. The emergence of vision language models has sparked an increase in demands …

De-identificationFairness

Improving VTE Identification through Adaptive NLP Model Selection and Clinical Expert Rule-based Classifier from Radiology Reports

2023-09-21 · Jamie Deng, Yusen Wu, Hilary Hayssen, Brain Englum 외

Rapid and accurate identification of Venous thromboembolism (VTE), a severe cardiovascular condition including deep vein thrombosis (DVT) and pulmonary embolism (PE), is important for effective treatment. Leveraging Natu…

Data AugmentationModel Selection

Self-Supervised Contextual Language Representation of Radiology Reports to Improve the Identification of Communication Urgency

2019-12-05 · Xing Meng, Craig H. Ganoe, Ryan T. Sieberg, Yvonne Y. Cheung 외

Machine learning methods have recently achieved high-performance in biomedical text analysis. However, a major bottleneck in the widespread application of these methods is obtaining the required large amounts of annotate…

Self-Supervised Learning

MiniGPT-Med: Large Language Model as a General Interface for Radiology Diagnosis

2024-07-04 · Asma Alkhaldi, Raneem Alnajim, Layan Alabdullatef, Rawan Alyahya 외

Recent advancements in artificial intelligence (AI) have precipitated significant breakthroughs in healthcare, particularly in refining diagnostic procedures. However, previous studies have often been constrained to limi…

DiagnosticLanguage ModelingLanguage ModellingLarge Language Model+4