paper-with-me

Papers

Learning Theory Can (Sometimes) Explain Generalisation in Graph Neural Networks

2021-12-07 · NeurIPS 2021 12 · Pascal Mattia Esser, Leena Chennuru Vankadara, Debarghya Ghoshdastidar

In recent years, several results in the supervised learning setting suggested that classical statistical learning-theoretic measures, such as VC dimension, do not adequately explain the performance of deep learning models which prompted a slew of work in the infinite-width and iteration regimes. However, there is little theoretical explanation for the success of neural networks beyond the supervised setting. In this paper we argue that, under some distributional assumptions, classical learning-theoretic measures can sufficiently explain generalization for graph neural networks in the transductive setting. In particular, we provide a rigorous analysis of the performance of neural networks in the context of transductive inference, specifically by analysing the generalisation properties of graph convolutional networks for the problem of node classification. While VC Dimension does result in trivial generalisation error bounds in this setting as well, we show that transductive Rademacher complexity can explain the generalisation properties of graph convolutional networks for stochastic block models. We further use the generalisation error bounds based on transductive Rademacher complexity to demonstrate the role of graph convolutions and network architectures in achieving smaller generalisation error and provide insights into when the graph structure can help in learning. The findings of this paper could re-new the interest in studying generalisation in neural networks in terms of learning-theoretic measures, albeit in specific problems.

📄 PDF Abstract BibTeX arXiv:2112.03968

Code (0)

등록된 구현이 없습니다.

Tasks

Learning TheoryNode Classification

Similar Papers 제목 키워드 기반

How Ensembles of Distilled Policies Improve Generalisation in Reinforcement Learning

2025-05-22 · Max Weltevrede, Moritz A. Zanger, Matthijs T. J. Spaan, Wendelin Böhmer

In the zero-shot policy transfer setting in reinforcement learning, the goal is to train an agent on a fixed set of training environments so that it can generalise to similar, but unseen, testing environments. Previous w…

Wasserstein PAC-Bayes Learning: Exploiting Optimisation Guarantees to Explain Generalisation

2023-04-14 · Maxime Haddouche, Benjamin Guedj

PAC-Bayes learning is an established framework to both assess the generalisation ability of learning algorithms, and design new learning algorithm by exploiting generalisation bounds as training objectives. Most of the e…

Exact Generalisation Error Exposes Benchmarks Skew Graph Neural Networks Success (or Failure)

2025-09-12 · Nil Ayday, Mahalakshmi Sabanayagam, Debarghya Ghoshdastidar arxiv

Graph Neural Networks (GNNs) have become the standard method for learning from networks across fields ranging from biology to social systems, yet a principled understanding of what enables them to extract meaningful repr…

Generalisation and the Geometry of Class Separability

2020-10-19 · NeurIPS Workshop DL-IG 2020 12 · Dominic Belcher, Adam Prugel-Bennett, Srinandan Dasmahapatra

Recent results in deep learning show that considering only the capacity of machines does not adequately explain the generalisation performance we can observe. We propose that by considering the geometry of the data we ca…

Deep Learning

Scale generalisation properties of extended scale-covariant and scale-invariant Gaussian derivative networks on image datasets with spatial scaling variations

2024-09-17 · Andrzej Perzanowski, Tony Lindeberg

This paper presents an in-depth analysis of the scale generalisation properties of the scale-covariant and scale-invariant Gaussian derivative networks, complemented with both conceptual and algorithmic extensions. For t…

Scale Generalisation