Latent Causal Invariant Model
Current supervised learning can learn spurious correlation during the data-fitting process, imposing issues regarding interpretability, out-of-distribution (OOD) generalization, and robustness. To avoid spurious correlation, we propose a Latent Causal Invariance Model (LaCIM) which pursues causal prediction. Specifically, we introduce latent variables that are separated into (a) output-causative factors and (b) others that are spuriously correlated to the output via confounders, to model the underlying causal factors. We further assume the generating mechanisms from latent space to observed data to be causally invariant. We give the identifiable claim of such invariance, particularly the disentanglement of output-causative factors from others, as a theoretical guarantee for precise inference and avoiding spurious correlation. We propose a Variational-Bayesian-based method for estimation and to optimize over the latent space for prediction. The utility of our approach is verified by improved interpretability, prediction power on various OOD scenarios (including healthcare) and robustness on security.
Code (0)
등록된 구현이 없습니다.
Tasks
DisentanglementmodelPredictionSimilar Papers 제목 키워드 기반
Learning Invariant Causal Mechanism from Vision-Language Models
Large-scale pre-trained vision-language models such as CLIP have been widely applied to a variety of downstream scenarios. In real-world applications, the CLIP model is often utilized in more diverse scenarios than those…
Causal InferenceDecision MakingTime Series Domain Adaptation via Latent Invariant Causal Mechanism
Time series domain adaptation aims to transfer the complex temporal dependence from the labeled source domain to the unlabeled target domain. Recent advances leverage the stable causal mechanism over observed variables t…
Domain AdaptationTime SeriesTime Series ClassificationVariational InferenceInvarGC: Invariant Granger Causality for Heterogeneous Interventional Time Series under Latent Confounding
Granger causality is widely used for causal structure discovery in complex systems from multivariate time series data. Traditional Granger causality tests based on linear models often fail to detect even mild non-linear …
Self-Supervised Learning with Data Augmentations Provably Isolates Content from Style
Self-supervised representation learning has shown remarkable success in a number of domains. A common practice is to perform data augmentation via hand-crafted transformations intended to leave the semantics of the data …
Data AugmentationDisentanglementImage ClassificationRepresentation Learning+1Causal Intervention for Subject-Deconfounded Facial Action Unit Recognition
Subject-invariant facial action unit (AU) recognition remains challenging for the reason that the data distribution varies among subjects. In this paper, we propose a causal inference framework for subject-invariant faci…
Causal InferenceFacial Action Unit Detection