paper-with-me

홈 › Papers

Out-of-Distribution Generalization via Risk Extrapolation (REx)

2020-03-02 · David Krueger, Ethan Caballero, Joern-Henrik Jacobsen, Amy Zhang, Jonathan Binas, Dinghuai Zhang, Remi Le Priol, Aaron Courville

Distributional shift is one of the major obstacles when transferring machine learning prediction systems from the lab to the real world. To tackle this problem, we assume that variation across training domains is representative of the variation we might encounter at test time, but also that shifts at test time may be more extreme in magnitude. In particular, we show that reducing differences in risk across training domains can reduce a model's sensitivity to a wide range of extreme distributional shifts, including the challenging setting where the input contains both causal and anti-causal elements. We motivate this approach, Risk Extrapolation (REx), as a form of robust optimization over a perturbation set of extrapolated domains (MM-REx), and propose a penalty on the variance of training risks (V-REx) as a simpler variant. We prove that variants of REx can recover the causal mechanisms of the targets, while also providing some robustness to changes in the input distribution ("covariate shift"). By appropriately trading-off robustness to causally induced distributional shifts and covariate shift, REx is able to outperform alternative methods such as Invariant Risk Minimization in situations where these types of shift co-occur.

📄 PDF Abstract BibTeX arXiv:2003.00688

Code (4)

capybaralet/REx_code_release 공식 구현 pytorch
facebookresearch/DomainBed pytorch
lingxiaoyuan/ood_mechanics pytorch
thuml/Transfer-Learning-Library pytorch

Tasks

Domain GeneralizationImage ClassificationOut-of-Distribution Generalization

Similar Papers 제목 키워드 기반

An Online Learning Approach to Interpolation and Extrapolation in Domain Generalization

2021-02-25 · Elan Rosenfeld, Pradeep Ravikumar, Andrej Risteski

A popular assumption for out-of-distribution generalization is that the training data comprises sub-datasets, each drawn from a distinct distribution; the goal is then to "interpolate" these distributions and "extrapolat…

Domain GeneralizationOut-of-Distribution Generalization

Robust Invariant Representation Learning by Distribution Extrapolation

2025-05-22 · Kotaro Yoshida, Slavakis Konstantinos

Invariant risk minimization (IRM) aims to enable out-of-distribution (OOD) generalization in deep learning by learning invariant representations. As IRM poses an inherently challenging bi-level optimization problem, most…

DiversityRepresentation Learning

Rule Extrapolation in Language Models: A Study of Compositional Generalization on OOD Prompts

2024-09-09 · Anna Mészáros, Szilvia Ujváry, Wieland Brendel, Patrik Reizinger 외

LLMs show remarkable emergent abilities, such as inferring concepts from presumably out-of-distribution prompts, known as in-context learning. Though this success is often attributed to the Transformer architecture, our …

In-Context LearningState Space Models

Risk Variance Penalization

2020-06-13 · Chuanlong Xie, Haotian Ye, Fei Chen, Yue Liu 외

The key of the out-of-distribution (OOD) generalization is to generalize invariance from training domains to target domains. The variance risk extrapolation (V-REx) is a practical OOD method, which depends on a domain-le…

Emputation: Identification-Guided Neural Imputation Framework

2026-07-06 · Yanjiao Yang, Yikun Zhang, Xinwei Shen, Yen-Chi Chen arxiv

We propose Emputation, a deep generative framework for learning imputation models. Emputation targets the extrapolation distribution of missing variables given observed variables, and training is guided by specific missi…