paper-with-me

Papers

Mechanisms for Hiding Sensitive Genotypes with Information-Theoretic Privacy

2020-07-10 · Fangwei Ye, Hyunghoon Cho, Salim El Rouayheb

Motivated by the growing availability of personal genomics services, we study an information-theoretic privacy problem that arises when sharing genomic data: a user wants to share his or her genome sequence while keeping the genotypes at certain positions hidden, which could otherwise reveal critical health-related information. A straightforward solution of erasing (masking) the chosen genotypes does not ensure privacy, because the correlation between nearby positions can leak the masked genotypes. We introduce an erasure-based privacy mechanism with perfect information-theoretic privacy, whereby the released sequence is statistically independent of the sensitive genotypes. Our mechanism can be interpreted as a locally-optimal greedy algorithm for a given processing order of sequence positions, where utility is measured by the number of positions released without erasure. We show that finding an optimal order is NP-hard in general and provide an upper bound on the optimal utility. For sequences from hidden Markov models, a standard modeling approach in genetics, we propose an efficient algorithmic implementation of our mechanism with complexity polynomial in sequence length. Moreover, we illustrate the robustness of the mechanism by bounding the privacy leakage from erroneous prior distributions. Our work is a step towards more rigorous control of privacy in genomic data sharing.

📄 PDF Abstract BibTeX arXiv:2007.05139

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Wasserstein Fair Classification

2019-07-28 · Ray Jiang, Aldo Pacchiano, Tom Stepleton, Heinrich Jiang 외

We propose an approach to fair classification that enforces independence between the classifier outputs and sensitive information by minimizing Wasserstein-1 distances. The approach has desirable theoretical properties a…

ClassificationFairnessGeneral Classification

Preventing Disclosure of Sensitive Knowledge by Hiding Inference

2013-08-28 · A. S. Syed Navaz, M. Ravi, T. Prabhu

Data Mining is a way of extracting data or uncovering hidden patterns of information from databases. So, there is a need to prevent the inference rules from being disclosed such that the more secure data sets cannot be i…

Data set operations to hide decision tree rules

2017-06-18 · Dimitris Kalles, Vassilios S. Verykios, Georgios Feretzakis, Athanasios Papagelis

This paper focuses on preserving the privacy of sensitive patterns when inducing decision trees. We adopt a record augmentation approach for hiding sensitive classification rules in binary datasets. Such a hiding methodo…

General Classification

Training-Free Coverless Multi-Image Steganography with Access Control

2026-03-10 · Minyeol Bae, Si-Hyeon Lee arxiv

Coverless Image Steganography (CIS) hides information without explicitly modifying a cover image, providing strong imperceptibility and inherent robustness to steganalysis. However, existing CIS methods largely lack robu…

Concealment of Intent: A Game-Theoretic Analysis

2025-05-27 · Xinbo Wu, Abhishek Umrawal, Lav R. Varshney

As large language models (LLMs) grow more capable, concerns about their safe deployment have also grown. Although alignment mechanisms have been introduced to deter misuse, they remain vulnerable to carefully designed ad…