paper-with-me

홈 › Papers

Privacy-Preserving Statistical Data Generation: Application to Sepsis Detection

2024-04-25 · Eric Macias-Fassio, Aythami Morales, Cristina Pruenza, Julian Fierrez

The biomedical field is among the sectors most impacted by the increasing regulation of Artificial Intelligence (AI) and data protection legislation, given the sensitivity of patient information. However, the rise of synthetic data generation methods offers a promising opportunity for data-driven technologies. In this study, we propose a statistical approach for synthetic data generation applicable in classification problems. We assess the utility and privacy implications of synthetic data generated by Kernel Density Estimator and K-Nearest Neighbors sampling (KDE-KNN) within a real-world context, specifically focusing on its application in sepsis detection. The detection of sepsis is a critical challenge in clinical practice due to its rapid progression and potentially life-threatening consequences. Moreover, we emphasize the benefits of KDE-KNN compared to current synthetic data generation methodologies. Additionally, our study examines the effects of incorporating synthetic data into model training procedures. This investigation provides valuable insights into the effectiveness of synthetic data generation techniques in mitigating regulatory constraints within the biomedical field.

📄 PDF Abstract BibTeX arXiv:2404.16638

Code (0)

등록된 구현이 없습니다.

Tasks

Privacy PreservingSynthetic Data Generation

Similar Papers 제목 키워드 기반

Privacy-Preserving Personalized Federated Learning for Distributed Photovoltaic Disaggregation under Statistical Heterogeneity

2025-04-25 · Xiaolu Chen, Chenghao Huang, Yanru Zhang, Hao Wang

The rapid expansion of distributed photovoltaic (PV) installations worldwide, many being behind-the-meter systems, has significantly challenged energy management and grid operations, as unobservable PV generation further…

energy managementFederated LearningPersonalized Federated LearningPrivacy Preserving

The Cost of Privacy: Optimal Rates of Convergence for Parameter Estimation with Differential Privacy

2019-02-12 · T. Tony Cai, Yichen Wang, Linjun Zhang

Privacy-preserving data analysis is a rising challenge in contemporary statistics, as the privacy guarantees of statistical methods are often achieved at the expense of accuracy. In this paper, we investigate the tradeof…

parameter estimationPrivacy Preservingregression

Synthetic Survival Data Generation for Heart Failure Prognosis Using Deep Generative Models

2025-09-04 · Chanon Puttanawarut, Natcha Fongsrisin, Porntep Amornritvanich, Panu Looareesuwan 외 arxiv

Background: Heart failure (HF) research is constrained by limited access to large, shareable datasets due to privacy regulations and institutional barriers. Synthetic data generation offers a promising solution to overco…

Synthetic Data Generation

Efficient Sparse Least Absolute Deviation Regression with Differential Privacy

2024-01-02 · Weidong Liu, Xiaojun Mao, Xiaofei Zhang, Xin Zhang

In recent years, privacy-preserving machine learning algorithms have attracted increasing attention because of their important applications in many scientific fields. However, in the literature, most privacy-preserving a…

Privacy Preservingregression

New Money: A Systematic Review of Synthetic Data Generation for Finance

2025-10-30 · James Meldrum, Basem Suleiman, Fethi Rabhi, Muhammad Johan Alibasa arxiv

Synthetic data generation has emerged as a promising approach to address the challenges of using sensitive financial data in machine learning applications. By leveraging generative models, such as Generative Adversarial …

Synthetic Data Generation