paper-with-me

Papers

Width-Robust Learnability in Mean-Field Bayesian Neural Networks

2026-07-07 · Dmitry Vaintrob, Kaarel Hänni arxiv

Infinite-width limits are a standard way to reason about neural networks, but it is not automatic that the limiting learner has the same complexity-theoretic inductive bias as large finite networks. We study this question for Bayesian neural networks at the mean-field, or critical feature-learning, scaling. The central quantity is the \emph{reduced entropy} \[ s_\infty(y,\varepsilon)=\limsup_N -\frac{1}{N}\log π_N^0(L\le \varepsilon), \] the intensive prior cost of representing a target function $y$ to population mean-squared error $\varepsilon$. Our main result is a width-robust learnability theorem. At fixed depth, a family of Boolean-cube targets is learnable from polynomially many samples at infinite width if and only if it is learnable at polynomial width, if and only if its reduced entropy is polynomially bounded. Equivalently, up to polynomial slack in accuracy, the Bayesian mean-field learner generalizes exactly on the targets that can be represented by polynomial-size networks. The forward direction is proved by a form of subsampling: from the infinitely many hidden neurons in the mean-field solution, one can select polynomially many representatives and still preserve the learned function on every input simultaneously. At the critical scaling this subsampling has both an `active'' component, which keeps the data-dependent low-dimensional statistics, and a `lazy'' component, which resamples the entropy-dominated directions from the prior. Thus the infinite-width mean-field limit gives a clean analytic description of learning without introducing spurious width-dependent generalization power.

📄 PDF Abstract BibTeX arXiv:2607.05735

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Wide Mean-Field Bayesian Neural Networks Ignore the Data

2022-02-23 · Beau Coker, Wessel P. Bruinsma, David R. Burt, Weiwei Pan 외

Bayesian neural networks (BNNs) combine the expressive power of deep learning with the advantages of Bayesian formalism. In recent years, the analysis of wide, deep BNNs has provided theoretical insight into their priors…

Variational Inference

Unified field theoretical approach to deep and recurrent neuronal networks

2021-12-10 · Kai Segadlo, Bastian Epping, Alexander van Meegen, David Dahmen 외

Understanding capabilities and limitations of different network architectures is of fundamental importance to machine learning. Bayesian inference on Gaussian processes has proven to be a viable approach for studying rec…

Bayesian InferenceGaussian Processes

A simple mean field model of feature learning

2025-10-16 · Niclas Göring, Chris Mingard, Yoonsoo Nam, Ard Louis arxiv

Feature learning (FL), where neural networks adapt their internal representations during training, remains poorly understood. Using methods from statistical physics, we derive a tractable, self-consistent mean-field (MF)…

Learning from almost nothing: How neural networks survive heavy input corruption

2026-06-09 · Justin Tahmassebpur, Asadullah Bhuiyan, Hyejin Kim, Omri Lesser arxiv

Learning from imperfect data is a central theme in machine learning, connecting practical questions of robustness to fundamental questions of learnability. Here we examine attribute noise: learning from corrupted inputs …

Multi-Item Mechanisms without Item-Independence: Learnability via Robustness

2019-11-06 · Johaness Brustle, Yang Cai, Constantinos Daskalakis

We study the sample complexity of learning revenue-optimal multi-item auctions. We obtain the first set of positive results that go beyond the standard but unrealistic setting of item-independence. In particular, we cons…