paper-with-me

Papers

Posterior Concentration for Sparse Deep Learning

2018-03-24 · NeurIPS 2018 12 · Nicholas Polson, Veronika Rockova

Spike-and-Slab Deep Learning (SS-DL) is a fully Bayesian alternative to Dropout for improving generalizability of deep ReLU networks. This new type of regularization enables provable recovery of smooth input-output maps with unknown levels of smoothness. Indeed, we show that the posterior distribution concentrates at the near minimax rate for $\alpha$-H\"older smooth maps, performing as well as if we knew the smoothness level $\alpha$ ahead of time. Our result sheds light on architecture design for deep neural networks, namely the choice of depth, width and sparsity level. These network attributes typically depend on unknown smoothness in order to be optimal. We obviate this constraint with the fully Bayes construction. As an aside, we show that SS-DL does not overfit in the sense that the posterior concentrates on smaller networks with fewer (up to the optimal number of) nodes and links. Our results provide new theoretical justifications for deep ReLU networks from a Bayesian point of view.

📄 PDF Abstract BibTeX arXiv:1803.09138

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Learning

Methods 이 논문이 사용한 방법론

ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…

Similar Papers 제목 키워드 기반

A Bayesian sparse factor model with adaptive posterior concentration

2023-05-29 · Ilsang Ohn, Lizhen Lin, Yongdai Kim

In this paper, we propose a new Bayesian inference method for a high-dimensional sparse factor model that allows both the factor dimensionality and the sparse structure of the loading matrix to be inferred. The novelty i…

Bayesian Inference

Posterior concentrations of fully-connected Bayesian neural networks with general priors on the weights

2024-03-21 · Insung Kong, Yongdai Kim

Bayesian approaches for training deep neural networks (BNNs) have received significant interest and have been effectively utilized in a wide range of applications. There have been several studies on the properties of pos…

Adaptive posterior concentration rates for sparse high-dimensional linear regression with random design and unknown error variance

2024-05-29 · The Tien Mai

This paper investigates sparse high-dimensional linear regression, particularly examining the properties of the posterior under conditions of random design and unknown error variance. We provide consistency results for t…

parameter estimationregression

Concentration of a sparse Bayesian model with Horseshoe prior in estimating high-dimensional precision matrix

2024-06-20 · The Tien Mai

Precision matrices are crucial in many fields such as social networks, neuroscience, and economics, representing the edge structure of Gaussian graphical models (GGMs), where a zero in an off-diagonal position of the pre…

Masked Bayesian Neural Networks : Theoretical Guarantee and its Posterior Inference

2023-05-24 · Insung Kong, Dongyoon Yang, Jongjin Lee, Ilsang Ohn 외

Bayesian approaches for learning deep neural networks (BNN) have been received much attention and successfully applied to various applications. Particularly, BNNs have the merit of having better generalization ability as…

Bayesian InferenceUncertainty Quantification