paper-with-me

홈 › Papers

CaSPECT: Discovering Causally Homogeneous Subgroups via Directed Spectral Clustering

2026-07-03 · Arghya Pratihar, Shinjon Chakraborty, Swagatam Das arxiv

We propose \textbf{CaSPECT}, a causal spectral clustering framework for discovering causally homogeneous subgroups from observational data. Rather than clustering in covariate space, CaSPECT defines similarity through the topology of a learned directed acyclic graph (DAG); a bootstrap-stabilised PC algorithm recovers the causal skeleton; a novel \emph{Orientation Validation Score} (OVS) combines PC bootstrap evidence with DirectLiNGAM to orient edges robustly; directed edges are weighted by backdoor-identified average treatment effects estimated via OLS or double machine learning. Chung's directed Laplacian provides a spectral embedding in which individuals close together share the same causal propagation pathways. We establish almost-sure consistency of the full pipeline and validate the method through a controlled simulation study and on LaLonde CPS1, IHDP, and 401(k) datasets, where CaSPECT recovers a positive and statistically significant treatment effect within the causally comparable subpopulation and corrects for severe confounding without requiring a pre-specified propensity score model.

📄 PDF Abstract BibTeX arXiv:2607.03364

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Automatic Discovery of Disease Subgroups by Contrasting with Healthy Controls

2026-05-20 · Robin Louiset, Edouard Duchesnay, Benoit Dufumier, Antoine Grigis 외 arxiv

In biomedical Subgroup Discovery, practitioners are interested in discovering interpretable and homogeneous subgroups within a group of patients. In this paper, assuming that healthy subjects (i.e., controls) share commo…

Discover and Mitigate Multiple Biased Subgroups in Image Classifiers

2024-03-19 · CVPR 2024 1 · Zeliang Zhang, Mingqian Feng, Zhiheng Li, Chenliang Xu

Machine learning models can perform well on in-distribution data but often fail on biased subgroups that are underrepresented in the training data, hindering the robustness of models for reliable applications. Such subgr…

Dimensionality ReductionSubgroup Discovery

FairVis: Visual Analytics for Discovering Intersectional Bias in Machine Learning

2019-04-10 · Ángel Alexander Cabrera, Will Epperson, Fred Hohman, Minsuk Kahng 외

The growing capability and accessibility of machine learning has led to its application to many real-world domains and data about people. Despite the benefits algorithmic systems may bring, models can reflect, inject, or…

BIG-bench Machine LearningFairnessSubgroup Discovery

Discovering Subgroups with Exceptional Survival Characteristics

2026-02-25 · Mhd Jawad Al Rahwanji, Sascha Xu, Nils Philipp Walter, Jilles Vreeken arxiv

In many applications, it is important to identify subpopulations that survive longer or shorter than the rest of the population. In medicine, for example, it allows determining which patients benefit from treatment, and …

T-Phenotype: Discovering Phenotypes of Predictive Temporal Patterns in Disease Progression

2023-02-24 · Yuchao Qin, Mihaela van der Schaar, Changhee Lee

Clustering time-series data in healthcare is crucial for clinical phenotyping to understand patients' disease progression patterns and to design treatment guidelines tailored to homogeneous patient subgroups. While rich …

ClusteringRepresentation LearningTime SeriesTime Series Analysis