paper-with-me

홈 › Papers

Categorical Co-Frequency Analysis: Clustering Diagnosis Codes to Predict Hospital Readmissions

2019-09-01 · Hallee E. Wong, Brianna C. Heggeseth, Steven J. Miller

Accurately predicting patients' risk of 30-day hospital readmission would enable hospitals to efficiently allocate resource-intensive interventions. We develop a new method, Categorical Co-Frequency Analysis (CoFA), for clustering diagnosis codes from the International Classification of Diseases (ICD) according to the similarity in relationships between covariates and readmission risk. CoFA measures the similarity between diagnoses by the frequency with which two diagnoses are split in the same direction versus split apart in random forests to predict readmission risk. Applying CoFA to de-identified data from Berkshire Medical Center, we identified three groups of diagnoses that vary in readmission risk. To evaluate CoFA, we compared readmission risk models using ICD majors and CoFA groups to a baseline model without diagnosis variables. We found substituting ICD majors for the CoFA-identified clusters simplified the model without compromising the accuracy of predictions. Fitting separate models for each ICD major and CoFA group did not improve predictions, suggesting that readmission risk may be more homogeneous that heterogeneous across diagnosis groups.

📄 PDF Abstract BibTeX arXiv:1909.00306

Code (1)

halleewong/cofa 공식 구현

Tasks

Clustering

Similar Papers 제목 키워드 기반

Categorical Unsupervised Variational Acoustic Clustering

2025-04-10 · Luan Vinícius Fiorio, Ivana Nikoloska, Ronald M. Aarts

We propose a categorical approach for unsupervised variational acoustic clustering of audio data in the time-frequency domain. The consideration of a categorical distribution enforces sharper clustering even when data po…

Clustering

K-Metamodes: frequency- and ensemble-based distributed k-modes clustering for security analytics

2019-09-30 · Andrey Sapegin, Christoph Meinel

Nowadays processing of Big Security Data, such as log messages, is commonly used for intrusion detection purposed. Its heterogeneous nature, as well as combination of numerical and categorical attributes does not allow t…

ClusteringIntrusion Detectionvalid

Co-clustering based exploratory analysis of mixed-type data tables

2022-12-22 · Aichetou Bouchareb, Marc Boullé, Fabrice Clérot, Fabrice Rossi

Co-clustering is a class of unsupervised data analysis techniques that extract the existing underlying dependency structure between the instances and variables of a data table as homogeneous blocks. Most of those techniq…

ClusteringVocal Bursts Type Prediction

Contrastive Representation Disentanglement for Clustering

2023-06-08 · Fei Ding, Dan Zhang, Yin Yang, Venkat Krovi 외

Clustering continues to be a significant and challenging task. Recent studies have demonstrated impressive results by applying clustering to feature representations acquired through self-supervised learning, particularly…

ClusteringContrastive LearningDisentanglementRepresentation Learning+1

Categorical data clustering: 25 years beyond K-modes

2024-08-30 · Tai Dinh, Wong Hauchi, Philippe Fournier-Viger, Daniil Lisik 외

The clustering of categorical data is a common and important task in computer science, offering profound implications across a spectrum of applications. Unlike purely numerical data, categorical data often lack inherent …

Categorical data clusteringClustering