paper-with-me

Papers

How do we measure privacy in text? A survey of text anonymization metrics

2025-11-30 · Yaxuan Ren, Krithika Ramesh, Yaxing Yao, Anjalie Field arxiv

In this work, we aim to clarify and reconcile metrics for evaluating privacy protection in text through a systematic survey. Although text anonymization is essential for enabling NLP research and model development in domains with sensitive data, evaluating whether anonymization methods sufficiently protect privacy remains an open challenge. In manually reviewing 47 papers that report privacy metrics, we identify and compare six distinct privacy notions, and analyze how the associated metrics capture different aspects of privacy risk. We then assess how well these notions align with legal privacy standards (HIPAA and GDPR), as well as user-centered expectations grounded in HCI studies. Our analysis offers practical guidance on navigating the landscape of privacy evaluation approaches further and highlights gaps in current practices. Ultimately, we aim to facilitate more robust, comparable, and legally aware privacy evaluations in text anonymization.

📄 PDF Abstract BibTeX arXiv:2512.01109

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Privacy Protection, Measurement Error, and the Integration of Remote Sensing and Socioeconomic Survey Data

2022-02-10 · Jeffrey D. Michler, Anna Josephson, Talip Kilic, Siobhan Murray

When publishing socioeconomic survey data, survey programs implement a variety of statistical methods designed to preserve privacy but which come at the cost of distorting the data. We explore the extent to which spatial…

Survey

A Survey on Current Trends and Recent Advances in Text Anonymization

2025-08-29 · Tobias Deußer, Lorenz Sparrenberg, Armin Berger, Max Hahnbück 외 arxiv

The proliferation of textual data containing sensitive personal information across various domains requires robust anonymization techniques to protect privacy and comply with regulations, while preserving data usability …

Large Language Models are Advanced Anonymizers

2024-02-21 · Robin Staab, Mark Vero, Mislav Balunović, Martin Vechev

Recent privacy research on large language models (LLMs) has shown that they achieve near-human-level performance at inferring personal data from online texts. With ever-increasing model capabilities, existing text anonym…

Text Anonymization

Adaptive Text Anonymization: Learning Privacy-Utility Trade-offs via Prompt Optimization

2026-02-24 · Gabriel Loiseau, Damien Sileo, Damien Riquet, Maxime Meyer 외 arxiv

Anonymizing textual documents is a highly context-sensitive problem: the appropriate balance between privacy protection and utility preservation varies with the data domain, privacy objectives, and downstream application…

The Text Anonymization Benchmark (TAB): A Dedicated Corpus and Evaluation Framework for Text Anonymization

2022-01-25 · Ildikó Pilán, Pierre Lison, Lilja Øvrelid, Anthi Papadopoulou 외

We present a novel benchmark and associated evaluation metrics for assessing the performance of text anonymization methods. Text anonymization, defined as the task of editing a text document to prevent the disclosure of …

De-identificationText Anonymization