Neural Network Extrapolations with G-invariances from a Single Environment
Despite —or maybe because of— their astonishing capacity to fit data, neural networks are widely believed to be unable to extrapolate beyond training data distribution. This work shows that, for extrapolations based on transformation groups, a model’s inability to extrapolate is unrelated to its capacity. Rather, the shortcoming is inherited from a classical statistical learning hypothesis: Examples not explicitly observed with infinitely many training examples cannot be likely outcomes in the learner’s model. In order to endow neural networks with the ability to extrapolate over group transformations, we introduce a learning framework guided by a new learning hypothesis: Any invariance to transformation groups is mandatory even without evidence, unless the learner deems it inconsistent with the training data. Unlike existing invariance-driven methods for counterfactual inference, this framework allows extrapolations from a single environment. Finally, we introduce sequence and image extrapolation tasks that validate our framework and showcase the shortcomings of traditional approaches.
Code (0)
등록된 구현이 없습니다.
Tasks
counterfactualCounterfactual InferenceSimilar Papers 제목 키워드 기반
Neural Networks for Learning Counterfactual G-Invariances from Single Environments
Despite -- or maybe because of -- their astonishing capacity to fit data, neural networks are believed to have difficulties extrapolating beyond training data distribution. This work shows that, for extrapolations based …
counterfactualOn Single-environment Extrapolations in Graph Classification and Regression Tasks
Extrapolation in graph classification/regression remains an underexplored area of an otherwise rapidly developing field. Our work contributes to a growing literature by providing the first systematic counterfactual model…
ClassificationcounterfactualGeneral ClassificationGraph Classification+1Unifying Variational Inference and PAC-Bayes for Supervised Learning that Scales
Neural Network based controllers hold enormous potential to learn complex, high-dimensional functions. However, they are prone to overfitting and unwarranted extrapolations. PAC Bayes is a generalized framework which is …
MuJoCoVariational InferenceMinimally Naturalistic Artificial Intelligence
The rapid advancement of machine learning techniques has re-energized research into general artificial intelligence. While the idea of domain-agnostic meta-learning is appealing, this emerging field must come to terms wi…
Common Sense ReasoningInductive BiasMeta-LearningLearning Online Visual Invariances for Novel Objects via Supervised and Self-Supervised Training
Humans can identify objects following various spatial transformations such as scale and viewpoint. This extends to novel objects, after a single presentation at a single pose, sometimes referred to as online invariance. …
Data AugmentationTranslation