paper-with-me

홈 › Papers

Latent Covariate Shift: Unlocking Partial Identifiability for Multi-Source Domain Adaptation

2022-08-30 · Yuhang Liu, Zhen Zhang, Dong Gong, Mingming Gong, Biwei Huang, Anton Van Den Hengel, Kun Zhang, Javen Qinfeng Shi

Multi-source domain adaptation (MSDA) addresses the challenge of learning a label prediction function for an unlabeled target domain by leveraging both the labeled data from multiple source domains and the unlabeled data from the target domain. Conventional MSDA approaches often rely on covariate shift or conditional shift paradigms, which assume a consistent label distribution across domains. However, this assumption proves limiting in practical scenarios where label distributions do vary across domains, diminishing its applicability in real-world settings. For example, animals from different regions exhibit diverse characteristics due to varying diets and genetics. Motivated by this, we propose a novel paradigm called latent covariate shift (LCS), which introduces significantly greater variability and adaptability across domains. Notably, it provides a theoretical assurance for recovering the latent cause of the label variable, which we refer to as the latent content variable. Within this new paradigm, we present an intricate causal generative model by introducing latent noises across domains, along with a latent content variable and a latent style variable to achieve more nuanced rendering of observational data. We demonstrate that the latent content variable can be identified up to block identifiability due to its versatile yet distinct causal structure. We anchor our theoretical insights into a novel MSDA method, which learns the label distribution conditioned on the identifiable latent content variable, thereby accommodating more substantial distribution shifts. The proposed approach showcases exceptional performance and efficacy on both simulated and real-world datasets.

📄 PDF Abstract BibTeX arXiv:2208.14161

Code (0)

등록된 구현이 없습니다.

Tasks

Domain Adaptation

Methods 이 논문이 사용한 방법론

ICA _Independent component analysis (ICA) is a statistical and computational technique for revealing hidden factors that underlie sets of random variables, measurements, or…

Similar Papers 제목 키워드 기반

Identifiable Latent Neural Causal Models

2024-03-23 · Yuhang Liu, Zhen Zhang, Dong Gong, Mingming Gong 외

Causal representation learning seeks to uncover latent, high-level causal representations from low-level observed data. It is particularly good at predictions under unseen distribution shifts, because these shifts can ge…

Representation Learning

Synthetic Potential Outcomes and Causal Mixture Identifiability

2024-05-29 · Bijan Mazaheri, Chandler Squires, Caroline Uhler

Heterogeneous data from multiple populations, sub-groups, or sources is often represented as a ``mixture model'' with a single latent class influencing all of the observed covariates. Heterogeneity can be resolved at mul…

Causal Inferencecounterfactual

Scalable Out-of-distribution Robustness in the Presence of Unobserved Confounders

2024-11-29 · Parjanya Prashant, Seyedeh Baharan Khatami, Bruno Ribeiro, Babak Salimi

We consider the task of out-of-distribution (OOD) generalization, where the distribution shift is due to an unobserved confounder ($Z$) affecting both the covariates ($X$) and the labels ($Y$). In this setting, tradition…

Domain Adaptation

Identifiable Multimodal Causal Representation Learning under Partial Latent Sharing

2026-05-18 · Manal Benhamza, Marianne Clausel, Myriam Tami arxiv

Causal representation learning (CRL) seeks to uncover meaningful latent variables and their corresponding causal structure from high-dimensional observational data. Although its significance, CRL identifiability remains …

Representation Learning

Using Monotonicity Restrictions to Identify Models with Partially Latent Covariates

2021-01-14 · Minji Bang, Wayne Yuan Gao, Andrew Postlewaite, Holger Sieg

This paper develops a new method for identifying econometric models with partially latent covariates. Such data structures arise in industrial organization and labor economics settings where data are collected using an i…