Learning Optimal Features via Partial Invariance
Learning models that are robust to distribution shifts is a key concern in the context of their real-life applicability. Invariant Risk Minimization (IRM) is a popular framework that aims to learn robust models from multiple environments. The success of IRM requires an important assumption: the underlying causal mechanisms/features remain invariant across environments. When not satisfied, we show that IRM can over-constrain the predictor and to remedy this, we propose a relaxation via $\textit{partial invariance}$. In this work, we theoretically highlight the sub-optimality of IRM and then demonstrate how learning from a partition of training domains can help improve invariant models. Several experiments, conducted both in linear settings as well as with deep neural networks on tasks over both language and image data, allow us to verify our conclusions.
Code (1)
Tasks
Domain GeneralizationSimilar Papers 제목 키워드 기반
Partial Optimal Transport with Applications on Positive-Unlabeled Learning
Classical optimal transport problem seeks a transportation map that preserves the total mass betwenn two probability distributions, requiring their mass to be the same. This may be too restrictive in certain applications…
Partial Law Invariance and Risk Measures
We introduce the concept of partial law invariance, generalizing the concepts of law invariance and probabilistic sophistication widely used in decision theory, as well as statistical and financial applications. This new…
Decision MakingDecision Making Under UncertaintyManagementModel Metamers Reveal Invariances in Graph Neural Networks
In recent years, deep neural networks have been extensively employed in perceptual systems to learn representations endowed with invariances, aiming to emulate the invariance mechanisms observed in the human brain. Howev…
Data Augmentation: A Fourier Analysis Perspective
Data augmentation is a simple and model-agnostic approach for exploiting known invariances in learning problems. Given a group acting on the input space, one augments the training set with transformed copies of each samp…
Data AugmentationContinuous Invariance Learning
Invariance learning methods aim to learn invariant features in the hope that they generalize under distributional shifts. Although many tasks are naturally characterized by continuous domains, current invariance learning…
Cloud ComputingCPU