On the Generalization and Adaption Performance of Causal Models
Learning models that offer robust out-of-distribution generalization and fast adaptation is a key challenge in modern machine learning. Modelling causal structure into neural networks holds the promise to accomplish robust zero and few-shot adaptation. Recent advances in differentiable causal discovery have proposed to factorize the data generating process into a set of modules, i.e. one module for the conditional distribution of every variable where only causal parents are used as predictors. Such a modular decomposition of knowledge enables adaptation to distributions shifts by only updating a subset of parameters. In this work, we systematically study the generalization and adaption performance of such modular neural causal models by comparing it to monolithic models and structured models where the set of predictors is not constrained to causal parents. Our analysis shows that the modular neural causal models outperform other models on both zero and few-shot adaptation in low data regimes and offer robust generalization. We also found that the effects are more significant for sparser graphs as compared to denser graphs.
Code (0)
등록된 구현이 없습니다.
Tasks
Causal DiscoveryOut-of-Distribution GeneralizationSimilar Papers 제목 키워드 기반
Domain Generalization via Causal Adjustment for Cross-Domain Sentiment Analysis
Domain adaption has been widely adapted for cross-domain sentiment analysis to transfer knowledge from the source domain to the target domain. Whereas, most methods are proposed under the assumption that the target (test…
Domain AdaptationDomain GeneralizationSentiment AnalysisTowards Generalizable Reinforcement Learning via Causality-Guided Self-Adaptive Representations
General intelligence requires quick adaption across tasks. While existing reinforcement learning (RL) methods have made progress in generalization, they typically assume only distribution changes between source and targe…
Atari Gamesreinforcement-learningReinforcement LearningReinforcement Learning (RL)+1Semantic-aware Grad-GAN for Virtual-to-Real Urban Scene Adaption
Recent advances in vision tasks (e.g., segmentation) highly depend on the availability of large-scale real-world image annotations obtained by cumbersome human labors. Moreover, the perception performance often drops sig…
Domain AdaptationSegmentationSemantic SegmentationGeneralization of Fitness Exercise Recognition from Doppler Measurements by Domain-adaption and Few-Shot Learning
In previous works, a mobile application was developed using an unmodified commercial off-the-shelf smartphone to recognize whole-body exercises. The working principle was based on the ultrasound Doppler sensing with the …
Domain AdaptationFew-Shot LearningPureGaze: Purifying Gaze Feature for Generalizable Gaze Estimation
Gaze estimation methods learn eye gaze from facial features. However, among rich information in the facial image, real gaze-relevant features only correspond to subtle changes in eye region, while other gaze-irrelevant f…
Domain AdaptationDomain GeneralizationGaze Estimation