paper-with-me

홈 › Papers

MMM: Clustering Multivariate Longitudinal Mixed-type Data

2025-09-15 · Francesco Amato, Julien Jacques arxiv

Multivariate longitudinal data of mixed-type are increasingly collected in many science domains. However, algorithms to cluster this kind of data remain scarce, due to the challenge to simultaneously model the within- and between-time dependence structures for multivariate data of mixed kind. We introduce the Mixture of Mixed-Matrices (MMM) model: reorganizing the data in a three-way structure and assuming that the non-continuous variables are observations of underlying latent continuous variables, the model relies on a mixture of matrix-variate normal distributions to perform clustering in the latent dimension. The MMM model is thus able to handle continuous, ordinal, binary, nominal and count data and to concurrently model the heterogeneity, the association among the responses and the temporal dependence structure in a parsimonious way and without assuming conditional independence. The inference is carried out through an MCMC-EM algorithm, which is detailed. An evaluation of the model through synthetic data shows its inference abilities. A real-world application on financial data is presented.

📄 PDF Abstract BibTeX arXiv:2509.12166

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Sequential Dirichlet Process Mixtures of Multivariate Skew t-distributions for Model-based Clustering of Flow Cytometry Data

2017-02-14 · Boris P. Hejblum, Chariff Alkhassim, Raphael Gottardo, François Caron 외

Flow cytometry is a high-throughput technology used to quantify multiple surface and intracellular markers at the level of a single cell. This enables to identify cell sub-types, and to determine their relative proportio…

ClusteringModel Selection

Contrastive Representation Learning of Longitudinal Disease Trajectories on Temporal Graphs

2026-07-28 · Bastian Pfeifer arxiv

Understanding disease trajectories from longitudinal clinical data remains challenging due to complex temporal dynamics and heterogeneous patient cohorts. Here, we present a contrastive representation learning framework …

Representation LearningContrastive Learning

Causal Inference on Multivariate and Mixed-Type Data

2017-02-21 · Alexander Marx, Jilles Vreeken

Given data over the joint distribution of two random variables $X$ and $Y$, we consider the problem of inferring the most likely causal direction between $X$ and $Y$. In particular, we consider the general case where bot…

Causal InferenceVocal Bursts Type Prediction

Large language models as synthetic clinical experts to inform longitudinal rare-disease modeling

2026-08-17 · Clemens Schächter, Astrid Pechmann, Janbernd Kirschner, Jan Hasenauer 외 arxiv

Due to the limited amount of information, modeling longitudinal rare-disease data can benefit from integrating clinical knowledge. Yet, elicitation of expert knowledge and formalization for model fitting is challenging, …

Representation LearningClinical Knowledge

Random survival forests with multivariate longitudinal endogenous covariates

2022-08-11 · Anthony Devaux, Catherine Helmer, Robin Genuer, Cécile Proust-Lima

Predicting the individual risk of a clinical event using the complete patient history is still a major challenge for personalized medicine. Among the methods developed to compute individual dynamic predictions, the joint…