paper-with-me

홈 › Papers

Handling missing values in healthcare data: A systematic review of deep learning-based imputation techniques

2022-10-15 · Mingxuan Liu, Siqi Li, Han Yuan, Marcus Eng Hock Ong, Yilin Ning, Feng Xie, Seyed Ehsan Saffari, Victor Volovici, Bibhas Chakraborty, Nan Liu

Objective: The proper handling of missing values is critical to delivering reliable estimates and decisions, especially in high-stakes fields such as clinical research. The increasing diversity and complexity of data have led many researchers to develop deep learning (DL)-based imputation techniques. We conducted a systematic review to evaluate the use of these techniques, with a particular focus on data types, aiming to assist healthcare researchers from various disciplines in dealing with missing values. Methods: We searched five databases (MEDLINE, Web of Science, Embase, CINAHL, and Scopus) for articles published prior to August 2021 that applied DL-based models to imputation. We assessed selected publications from four perspectives: health data types, model backbone (i.e., main architecture), imputation strategies, and comparison with non-DL-based methods. Based on data types, we created an evidence map to illustrate the adoption of DL models. Results: We included 64 articles, of which tabular static (26.6%, 17/64) and temporal data (37.5%, 24/64) were the most frequently investigated. We found that model backbone(s) differed among data types as well as the imputation strategy. The "integrated" strategy, that is, the imputation task being solved concurrently with downstream tasks, was popular for tabular temporal (50%, 12/24) and multi-modal data (71.4%, 5/7), but limited for other data types. Moreover, DL-based imputation methods yielded better imputation accuracy in most studies, compared with non-DL-based methods. Conclusion: DL-based imputation models can be customized based on data type, addressing the corresponding missing patterns, and its associated "integrated" strategy can enhance the efficacy of imputation, especially in scenarios where data is complex. Future research may focus on the portability and fairness of DL-based models for healthcare data imputation.

📄 PDF Abstract BibTeX arXiv:2210.08258

Code (0)

등록된 구현이 없습니다.

Tasks

ArticlesFairnessImputationMissing Values

Similar Papers 제목 키워드 기반

A Hamiltonian Monte Carlo Model for Imputation and Augmentation of Healthcare Data

2021-03-03 · Narges Pourshahrokhi, Samaneh Kouchaki, Kord M. Kober, Christine Miaskowski 외

Missing values exist in nearly all clinical studies because data for a variable or question are not collected or not available. Inadequate handling of missing values can lead to biased results and loss of statistical pow…

Bayesian InferenceImputationMissing Values

Missing Values and Imputation in Healthcare Data: Can Interpretable Machine Learning Help?

2023-04-23 · Zhi Chen, Sarah Tan, Urszula Chajewska, Cynthia Rudin 외

Missing values are a fundamental problem in data science. Many datasets have missing values that must be properly handled because the way missing values are treated can have large impact on the resulting machine learning…

ImputationInterpretable Machine LearningMissing Values

On the Performance of Imputation Techniques for Missing Values on Healthcare Datasets

2024-03-13 · Luke Oluwaseye Joel, Wesley Doorsamy, Babu Sena Paul

Missing values or data is one popular characteristic of real-world datasets, especially healthcare data. This could be frustrating when using machine learning algorithms on such datasets, simply because most machine lear…

feature selectionImputationMissing Values

Impact of Missing Values in Machine Learning: A Comprehensive Analysis

2024-10-10 · Abu Fuad Ahmad, Md Shohel Sayeed, Khaznah Alshammari, Istiaque Ahmed

Machine learning (ML) has become a ubiquitous tool across various domains of data mining and big data analysis. The efficacy of ML models depends heavily on high-quality datasets, which are often complicated by the prese…

ImputationMissing ValuesModel Selection

Multilevel Weighted Support Vector Machine for Classification on Healthcare Data with Missing Values

2016-04-07 · Talayeh Razzaghi, Oleg Roderick, Ilya Safro, Nicholas Marko

This work is motivated by the needs of predictive analytics on healthcare data as represented by Electronic Medical Records. Such data is invariably problematic: noisy, with missing entries, with imbalance in classes of …

ClassificationGeneral ClassificationImputationMissing Values+1