paper-with-me

Papers

Statistical Jump Model for Mixed-Type Data with Missing Data Imputation

2024-09-02 · Federico P. Cortese, Antonio Pievatolo

In this paper, we address the challenge of clustering mixed-type data with temporal evolution by introducing the statistical jump model for mixed-type data. This novel framework incorporates regime persistence, enhancing interpretability and reducing the frequency of state switches, and efficiently handles missing data. The model is easily interpretable through its state-conditional means and modes, making it accessible to practitioners and policymakers. We validate our approach through extensive simulation studies and an empirical application to air quality data, demonstrating its superiority in inferring persistent air quality regimes compared to the traditional air quality index. Our contributions include a robust method for mixed-type temporal clustering, effective missing data management, and practical insights for environmental monitoring.

📄 PDF Abstract BibTeX arXiv:2409.01208

Code (1)

FedericoCortese/JM-mix 공식 구현

Tasks

ClusteringImputationManagement

Similar Papers 제목 키워드 기반

Precision Adaptive Imputation Network : An Unified Technique for Mixed Datasets

2025-01-18 · Harsh Joshi, Rajeshwari Mistri, Manasi Mali, Nachiket Kapure 외

The challenge of missing data remains a significant obstacle across various scientific domains, necessitating the development of advanced imputation techniques that can effectively address complex missingness patterns. T…

Imputation

Distances with mixed type variables some modified Gower's coefficients

2021-01-07 · Marcello D'Orazio

Nearest neighbor methods have become popular in official statistics, mainly in imputation or in statistical matching problems; they play a key role in machine learning too, where a high number of variants have been propo…

Density EstimationImputationMissing ValuesVocal Bursts Type Prediction

Statistical-Neural Interaction Networks for Interpretable Mixed-Type Data Imputation

2026-01-18 · Ou Deng, Shoji Nishimura, Atsushi Ogihara, Qun Jin arxiv

Real-world tabular databases routinely combine continuous measurements and categorical records, yet missing entries are pervasive and can distort downstream analysis. We propose Statistical-Neural Interaction (SNI), an i…

Deep Generalised Mixed Models: a Novel Neural Network Structure for Analysing Hierarchical Data

2026-08-06 · Nina van Gerwen, Dimitris Rizopoulos, Manon Hillegers, Loes Keijsers 외 arxiv

The experience sampling method (ESM) is a longitudinal research design where participants report their thoughts, emotional states and behaviours multiple times a day. Our work is motivated by such data collected by the G…

Data Augmentation

Model-based Clustering with Missing Not At Random Data

2021-12-20 · Aude Sportisse, Matthieu Marbac, Fabien Laporte, Gilles Celeux 외

Model-based unsupervised learning, as any learning task, stalls as soon as missing data occurs. This is even more true when the missing data are informative, or said missing not at random (MNAR). In this paper, we propos…

ClusteringImputation