paper-with-me

홈 › Papers

Large deviation principles for convolutional Bayesian neural networks

2026-03-06 · Federico Bassetti, Vassili De Palma, Lucia Ladelli arxiv

While suitably scaled CNNs with Gaussian initialization are known to converge to Gaussian processes as the number of channels diverges, little is known beyond this Gaussian limit. We establish a large deviation principle (LDP) for convolutional neural networks in the infinite-channel regime. We consider a broad class of multidimensional CNN architectures characterized by general receptive fields encoded through a patch-extractor function satisfying mild structural assumptions. Our main result establishes a large deviation principle (LDP) for the sequence of conditional covariance matrices under Gaussian prior distribution on the weights. We further derive an LDP for the posterior distribution obtained by conditioning on a finite number of observations. In addition, we provide a streamlined proof of the concentration of the conditional covariances and of the Gaussian equivalence of the network. To the best of our knowledge, this is the first large deviation principle established for convolutional neural networks.

📄 PDF Abstract BibTeX arXiv:2603.06023

Code (0)

등록된 구현이 없습니다.

Tasks

Gaussian Processes

Similar Papers 제목 키워드 기반

Uncertainty Estimation and Generalization Bounds for Modern Deep Learning

2026-06-11 · Luis A. Ortega arxiv

This thesis investigates how Bayesian principles can deepen our understanding of modern deep learning systems. While neural networks achieve remarkable predictive performance, their ability to generalize and to quantify …

Bayesian Inference

Stein Variational Gradient Descent: many-particle and long-time asymptotics

2021-02-25 · Nikolas Nüsken, D. R. Michiel Renger

Stein variational gradient descent (SVGD) refers to a class of methods for Bayesian inference based on interacting particle systems. In this paper, we consider the originally proposed deterministic dynamics as well as a …

Bayesian InferenceVariational Inference

Feature learning in finite-width Bayesian deep linear networks with multiple outputs and convolutional layers

2024-06-05 · Federico Bassetti, Marco Gherardi, Alessandro Ingrosso, Mauro Pastore 외

Deep linear networks have been extensively studied, as they provide simplified models of deep learning. However, little is known in the case of finite-width architectures with multiple outputs and convolutional layers. I…

Large deviation principles and evolutionary multiple structure alignment of non-coding RNA

2024-05-22 · Brandon Legried

Non-coding RNA are functional molecules that are not translated into proteins. Their function comes as important regulators of biological function. Because they are not translated, they need not be as stable as other typ…

Kernel Renormalization in Bayesian Deep Neural Networks: the Equivalent Wishart Ansatz in the Proportional Regime

2026-05-28 · Paolo Baglioni, Christian Keup, Vincenzo Zimbardo, Rosalba Pacelli 외 arxiv

The scaling limit where both the size of the training set $P$ and the width $N$ of a deep neural network grow at the same rate, the so-called proportional-width regime, has been intensely studied for shallow, single-hidd…

Representation Learning