Mind the gap: an experimental evaluation of imputation of missing values techniques in time series
Recording sensor data is seldom a perfect process. Failures in power, communication or storage can leave occasional blocks of data missing, affecting not only real-time monitoring but also compromising the quality of near- and off-line data analysis. Several recovery (imputation) algorithms have been proposed to replace missing blocks. Unfortunately, little is known about their relative performance, as existing comparisons are limited to either a small subset of relevant algorithms or to very few datasets or often both. Drawing general conclusions in this case remains a challenge. In this paper, we empirically compare twelve recovery algorithms using a novel benchmark. All but two of the algorithms were re-implemented in a uniform test environment. The benchmark gathers ten different datasets, which collectively represent a broad range of applications. Our benchmark allows us to fairly evaluate the strengths and weaknesses of each approach, and to recommend the best technique on a use-case basis. It also allows us to identify the limitations of the current body of algorithms and suggest future research directions.
Code (1)
Tasks
ImputationMissing ValuesTime SeriesTime Series AnalysisSimilar Papers 제목 키워드 기반
Generative Semi-supervised Learning for Multivariate Time Series Imputation
The missing values, widely existed in multivariate time series data, hinder the effective data analysis. Existing time series imputation methods do not make full use of the label information in real-life time series data…
Generative Adversarial NetworkImputationMissing ValuesMultivariate Time Series Imputation+2Explainability of Machine Learning Models under Missing Data
Missing data is a prevalent issue that can significantly impair model performance and explainability. This paper briefly summarizes the development of the field of missing data with respect to Explainable Artificial Inte…
Explainable artificial intelligenceFeature ImportanceImputationMissing ValuesIterative missing value imputation based on feature importance
Many datasets suffer from missing values due to various reasons,which not only increases the processing difficulty of related tasks but also reduces the accuracy of classification. To address this problem, the mainstream…
Feature ImportanceImputationMatrix CompletionMissing ValuesMissing Value Estimation using Clustering and Deep Learning within Multiple Imputation Framework
Missing values in tabular data restrict the use and performance of machine learning, requiring the imputation of missing values. The most popular imputation algorithm is arguably multiple imputations using chains of equa…
ClusteringEnsemble LearningImputationMissing ValuesDeep Imputation of Missing Values in Time Series Health Data: A Review with Benchmarking
The imputation of missing values in multivariate time series (MTS) data is critical in ensuring data quality and producing reliable data-driven predictive models. Apart from many statistical approaches, a few recent stud…
BenchmarkingDeep LearningImputationMissing Values+2