paper-with-me

홈 › Papers

A theory of independent mechanisms for extrapolation in generative models

2020-04-01 · Michel Besserve, Rémy Sun, Dominik Janzing, Bernhard Schölkopf

Generative models can be trained to emulate complex empirical data, but are they useful to make predictions in the context of previously unobserved environments? An intuitive idea to promote such extrapolation capabilities is to have the architecture of such model reflect a causal graph of the true data generating process, such that one can intervene on each node independently of the others. However, the nodes of this graph are usually unobserved, leading to overparameterization and lack of identifiability of the causal structure. We develop a theoretical framework to address this challenging situation by defining a weaker form of identifiability, based on the principle of independence of mechanisms. We demonstrate on toy examples that classical stochastic gradient descent can hinder the model's extrapolation capabilities, suggesting independence of mechanisms should be enforced explicitly during training. Experiments on deep generative models trained on real world data support these insights and illustrate how the extrapolation capabilities of such models can be leveraged.

📄 PDF Abstract BibTeX arXiv:2004.00184

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Extrapolation in Statistical Learning with Extreme Value Theory

2026-05-03 · Sebastian Engelke, Nicola Gnecco, Anne Sabourin arxiv

Extreme value theory provides rigorous theory and statistical tools for extrapolation in machine learning, particularly in settings where traditional methods struggle due to data scarcity in the tails. A broad range of t…

Anomaly Detection

Towards Visual Foundational Models of Physical Scenes

2023-06-06 · Chethan Parameshwara, Alessandro Achille, Xiaolong Li, Jiawei Mo 외

We describe a first step towards learning general-purpose visual representations of physical scenes using only image prediction as a training criterion. To do so, we first define "physical scene" and show that, even thou…

NeRF

Spiral Generative Network for Image Extrapolation

2020-08-01 · ECCV 2020 8 · Dongsheng Guo, Hongzhi Liu, Haoru Zhao, Yunhao Cheng 외

In this paper, motivated by human natural ability to perceive unseen surroundings imaginatively, we propose a novel Spiral Generative Network, SpiralNet, to perform image extrapolation in a spiral manner, which regards e…

From Identifiable Causal Representations to Controllable Counterfactual Generation: A Survey on Causal Generative Modeling

2023-10-17 · Aneesh Komanduri, Xintao Wu, Yongkai Wu, Feng Chen

Deep generative models have shown tremendous capability in data density estimation and data generation from finite samples. While these models have shown impressive performance by learning correlations among features in …

counterfactualDensity EstimationFairnessOut-of-Distribution Generalization+1

Towards Understanding Extrapolation: a Causal Lens

2025-01-15 · Lingjing Kong, Guangyi Chen, Petar Stojanov, Haoxuan Li 외

Canonical work handling distribution shifts typically necessitates an entire target distribution that lands inside the training distribution. However, practical scenarios often involve only a handful of target samples, p…