paper-with-me

Papers

Robust Domain Generalisation with Causal Invariant Bayesian Neural Networks

2024-10-08 · Gaël Gendron, Michael Witbrock, Gillian Dobbie

Deep neural networks can obtain impressive performance on various tasks under the assumption that their training domain is identical to their target domain. Performance can drop dramatically when this assumption does not hold. One explanation for this discrepancy is the presence of spurious domain-specific correlations in the training data that the network exploits. Causal mechanisms, in the other hand, can be made invariant under distribution changes as they allow disentangling the factors of distribution underlying the data generation. Yet, learning causal mechanisms to improve out-of-distribution generalisation remains an under-explored area. We propose a Bayesian neural architecture that disentangles the learning of the the data distribution from the inference process mechanisms. We show theoretically and experimentally that our model approximates reasoning under causal interventions. We demonstrate the performance of our method, outperforming point estimate-counterparts, on out-of-distribution image recognition tasks where the data distribution acts as strong adversarial confounders.

📄 PDF Abstract BibTeX arXiv:2410.06349

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Causal Inference via Style Transfer for Out-of-distribution Generalisation

2022-12-06 · Toan Nguyen, Kien Do, Duc Thanh Nguyen, Bao Duong 외

Out-of-distribution (OOD) generalisation aims to build a model that can generalise well on an unseen target domain using knowledge from multiple source domains. To this end, the model should seek the causal dependence be…

Causal InferenceImage ClassificationImage GenerationStyle Transfer

Can Large Language Models Learn Independent Causal Mechanisms?

2024-02-04 · Gaël Gendron, Bao Trung Nguyen, Alex Yuxuan Peng, Michael Witbrock 외

Despite impressive performance on language modelling and complex reasoning tasks, Large Language Models (LLMs) fall short on the same tasks in uncommon settings or with distribution shifts, exhibiting a lack of generalis…

Language Modelling

A Practical PAC-Bayes Generalisation Bound for Deep Learning

2021-09-29 · Diego Granziol, Mingtian Zhang, Nicholas Baskerville

Under a PAC-Bayesian framework, we derive an implementation efficient parameterisation invariant metric to measure the difference between our true and empirical risk. We show that for solutions of low training loss, this…

Deep Learning

Bayesian Hierarchical Invariant Prediction

2025-05-16 · Francisco Madaleno, Pernille Julie Viuff Sand, Francisco C. Pereira, Sergio Hernan Garrido Mejia

We propose Bayesian Hierarchical Invariant Prediction (BHIP) reframing Invariant Causal Prediction (ICP) through the lens of Hierarchical Bayes. We leverage the hierarchical structure to explicitly test invariance of cau…

Prediction

Lifted Causal Inference

2026-06-26 · Malte Luttermann, Tanya Braun, Ralf Möller, Marcel Gehrke arxiv

Lifted inference exploits indistinguishabilities in probabilistic graphical models by using a representative for indistinguishable objects, thereby speeding up query answering while maintaining exact answers. In this art…

Causal Inference