paper-with-me

홈 › Papers

Bayes without Underfitting: Fully Correlated Deep Learning Posteriors via Alternating Projections

2024-10-22 · Marco Miani, Hrittik Roy, Søren Hauberg

Bayesian deep learning all too often underfits so that the Bayesian prediction is less accurate than a simple point estimate. Uncertainty quantification then comes at the cost of accuracy. For linearized models, the null space of the generalized Gauss-Newton matrix corresponds to parameters that preserve the training predictions of the point estimate. We propose to build Bayesian approximations in this null space, thereby guaranteeing that the Bayesian predictive does not underfit. We suggest a matrix-free algorithm for projecting onto this null space, which scales linearly with the number of parameters and quadratically with the number of output dimensions. We further propose an approximation that only scales linearly with parameters to make the method applicable to generative models. An extensive empirical evaluation shows that the approach scales to large models, including vision transformers with 28 million parameters.

📄 PDF Abstract BibTeX arXiv:2410.16901

Code (0)

등록된 구현이 없습니다.

Tasks

Deep LearningUncertainty Quantification

Similar Papers 제목 키워드 기반

If there is no underfitting, there is no Cold Posterior Effect

2023-10-02 · Yijie Zhang, Yi-Shan Wu, Luis A. Ortega, Andrés R. Masegosa

The cold posterior effect (CPE) (Wenzel et al., 2020) in Bayesian deep learning shows that, for posteriors with a temperature $T<1$, the resulting posterior predictive could have better performances than the Bayesian pos…

On Cold Posteriors of Probabilistic Neural Networks: Understanding the Cold Posterior Effect and A New Way to Learn Cold Posteriors with Tight Generalization Guarantees

2024-10-20 · Yijie Zhang

Bayesian inference provides a principled probabilistic framework for quantifying uncertainty by updating beliefs based on prior knowledge and observed data through Bayes' theorem. In Bayesian deep learning, neural networ…

Bayesian InferenceGeneralization Bounds

Wide Mean-Field Bayesian Neural Networks Ignore the Data

2022-02-23 · Beau Coker, Wessel P. Bruinsma, David R. Burt, Weiwei Pan 외

Bayesian neural networks (BNNs) combine the expressive power of deep learning with the advantages of Bayesian formalism. In recent years, the analysis of wide, deep BNNs has provided theoretical insight into their priors…

Variational Inference

CATVI: Conditional and Adaptively Truncated Variational Inference for Hierarchical Bayesian Nonparametric Models

2020-01-13 · Yirui Liu, Xinghao Qiao, Jessica Lam

Current variational inference methods for hierarchical Bayesian nonparametric models can neither characterize the correlation structure among latent variables due to the mean-field setting, nor infer the true posterior d…

ClusteringTopic ModelsVariational Inference

Collapsed Variational Bounds for Bayesian Neural Networks

2021-12-01 · NeurIPS 2021 12 · Marcin Tomczak, Siddharth Swaroop, Andrew Foong, Richard Turner

Recent interest in learning large variational Bayesian Neural Networks (BNNs) has been partly hampered by poor predictive performance caused by underfitting, and their performance is known to be very sensitive to the pri…

Variational Inference