paper-with-me

Papers

Can Microcanonical Langevin Dynamics Leverage Mini-Batch Gradient Noise?

2026-02-06 · Emanuel Sommer, Kangning Diao, Jakob Robnik, Uros Seljak, David Rügamer arxiv

Scaling inference methods such as Markov chain Monte Carlo to high-dimensional models remains a central challenge in Bayesian deep learning. A promising recent proposal, microcanonical Langevin Monte Carlo, has shown state-of-the-art performance across a wide range of problems. However, its reliance on full-dataset gradients makes it prohibitively expensive for large-scale problems. This paper addresses a fundamental question: Can microcanonical dynamics effectively leverage mini-batch gradient noise? We provide the first systematic study of this problem, establishing a novel continuous-time theoretical analysis of stochastic-gradient microcanonical dynamics. We reveal two critical failure modes: a theoretically derived bias due to anisotropic gradient noise and numerical instabilities in complex high-dimensional posteriors. To tackle these issues, we propose a principled gradient noise preconditioning scheme shown to significantly reduce this bias and develop a novel, energy-variance-based adaptive tuner that automates step size selection and dynamically informs numerical guardrails. The resulting algorithm is a robust and scalable microcanonical Monte Carlo sampler that achieves state-of-the-art performance on challenging high-dimensional inference tasks like Bayesian neural networks. Combined with recent ensemble techniques, our work unlocks a new class of stochastic microcanonical Langevin ensemble (SMILE) samplers for large-scale Bayesian inference.

📄 PDF Abstract BibTeX arXiv:2602.06500

Code (0)

등록된 구현이 없습니다.

Tasks

Bayesian Inference

Similar Papers 제목 키워드 기반

Fluctuation without dissipation: Microcanonical Langevin Monte Carlo

2023-03-31 · Jakob Robnik, Uroš Seljak

Stochastic sampling algorithms such as Langevin Monte Carlo are inspired by physical systems in a heat bath. Their equilibrium distribution is the canonical ensemble given by a prescribed target distribution, so they mus…

Microcanonical Langevin Ensembles: Advancing the Sampling of Bayesian Neural Networks

2025-02-10 · Emanuel Sommer, Jakob Robnik, Giorgi Nozadze, Uros Seljak 외

Despite recent advances, sampling-based inference for Bayesian Neural Networks (BNNs) remains a significant challenge in probabilistic deep learning. While sampling-based approaches do not require a variational distribut…

NavigateProbabilistic Deep LearningUncertainty Quantification

Quantifying the mini-batching error in Bayesian inference for Adaptive Langevin dynamics

2021-05-21 · Inass Sekkat, Gabriel Stoltz

Bayesian inference allows to obtain useful information on the parameters of models, either in computational statistics or more recently in the context of Bayesian Neural Networks. The computational cost of usual Monte Ca…

Bayesian InferenceFriction

Sampling from Bayesian Neural Network Posteriors with Symmetric Minibatch Splitting Langevin Dynamics

2024-10-14 · Daniel Paulin, Peter A. Whalley, Neil K. Chada, Benedict Leimkuhler

We propose a scalable kinetic Langevin dynamics algorithm for sampling parameter spaces of big data and AI applications. Our scheme combines a symmetric forward/backward sweep over minibatches with a symmetric discretiza…

Generalized EXTRA stochastic gradient Langevin dynamics

2024-12-02 · Mert Gurbuzbalaban, Mohammad Rafiqul Islam, Xiaoyu Wang, Lingjiong Zhu

Langevin algorithms are popular Markov Chain Monte Carlo methods for Bayesian learning, particularly when the aim is to sample from the posterior distribution of a parametric model, given the input data and the prior dis…