paper-with-me

Papers

Provably Strict Generalisation Benefit for Equivariant Models

2021-02-20 · Bryn Elesedy, Sheheryar Zaidi

It is widely believed that engineering a model to be invariant/equivariant improves generalisation. Despite the growing popularity of this approach, a precise characterisation of the generalisation benefit is lacking. By considering the simplest case of linear models, this paper provides the first provably non-zero improvement in generalisation for invariant/equivariant models when the target distribution is invariant/equivariant with respect to a compact group. Moreover, our work reveals an interesting relationship between generalisation, the number of training examples and properties of the group action. Our results rest on an observation of the structure of function spaces under averaging operators which, along with its consequences for feature averaging, may be of independent interest.

📄 PDF Abstract BibTeX arXiv:2102.10333

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Provably Strict Generalisation Benefit for Invariance in Kernel Methods

2021-06-04 · NeurIPS 2021 12 · Bryn Elesedy

It is a commonly held belief that enforcing invariance improves generalisation. Although this approach enjoys widespread popularity, it is only very recently that a rigorous theoretical demonstration of this benefit has …

Equivariant MuZero

2023-02-09 · Andreea Deac, Théophane Weber, George Papamakarios

Deep reinforcement learning repeatedly succeeds in closed, well-defined domains such as games (Chess, Go, StarCraft). The next frontier is real-world scenarios, where setups are numerous and varied. For this, agents need…

Deep Reinforcement LearningModel-based Reinforcement Learningreinforcement-learningReinforcement Learning (RL)+1

Symmetry and Generalisation in Machine Learning

2025-01-07 · Hayder Elesedy

This work is about understanding the impact of invariance and equivariance on generalisation in supervised learning. We use the perspective afforded by an averaging operator to show that for any predictor that is not equ…

Inductive Biasregression

PRISM: Parallel Reward Integration with Symmetry for MORL

2026-02-20 · Finn van der Knaap, Kejiang Qian, Zheng Xu, Fengxiang He arxiv

This work studies heterogeneous Multi-Objective Reinforcement Learning (MORL), where objectives can differ sharply in temporal frequency. Such heterogeneity allows dense objectives to dominate learning, while sparse long…

Reinforcement Learning

Exact equivariance, kept through training, buys zero-shot generalisation across the symmetry group

2026-06-02 · Hongbo Wang arxiv

A latent world model built from an equivariant encoder and predictor inherits a provable symmetry of its training loss: when the dynamics carries a group $G$ acting on latents by an orthogonal representation $ρ(g)$, the …