paper-with-me

홈 › Papers

Langevin Monte-Carlo Provably Learns Depth Two Neural Nets at Any Size and Data

2025-03-13 · Dibyakanti Kumar, Samyak Jha, Anirbit Mukherjee

In this work, we will establish that the Langevin Monte-Carlo algorithm can learn depth-2 neural nets of any size and for any data and we give non-asymptotic convergence rates for it. We achieve this via showing that under Total Variation distance and q-Renyi divergence, the iterates of Langevin Monte Carlo converge to the Gibbs distribution of Frobenius norm regularized losses for any of these nets, when using smooth activations and in both classification and regression settings. Most critically, the amount of regularization needed for our results is independent of the size of the net. This result combines several recent observations, like our previous papers showing that two-layer neural loss functions can always be regularized by a certain constant amount such that they satisfy the Villani conditions, and thus their Gibbs measures satisfy a Poincare inequality.

📄 PDF Abstract BibTeX arXiv:2503.10428

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Subspace Langevin Monte Carlo

2024-12-18 · Tyler Maunu, Jiayi Yao

Sampling from high-dimensional distributions has wide applications in data science and machine learning but poses significant computational challenges. We introduce Subspace Langevin Monte Carlo (SLMC), a novel and effic…

Computational Efficiency

Regime-Switching Langevin Monte Carlo Algorithms

2025-08-31 · Xiaoyu Wang, Yingli Wang, Lingjiong Zhu arxiv

Langevin Monte Carlo (LMC) algorithms are popular Markov Chain Monte Carlo (MCMC) methods to sample a target probability distribution, which arises in many applications in machine learning. Inspired by regime-switching s…

Accelerating Langevin Monte Carlo Sampling: A Large Deviations Analysis

2025-03-24 · Nian Yao, Pervez Ali, Xihua Tao, Lingjiong Zhu

Langevin algorithms are popular Markov chain Monte Carlo methods that are often used to solve high-dimensional large-scale sampling problems in machine learning. The most classical Langevin Monte Carlo algorithm is based…

Bregman Proximal Langevin Monte Carlo via Bregman--Moreau Envelopes

2022-07-10 · Tim Tsz-Kit Lau, Han Liu

We propose efficient Langevin Monte Carlo algorithms for sampling distributions with nonsmooth convex composite potentials, which is the sum of a continuously differentiable function and a possibly nonsmooth function. We…

High-Order Langevin Monte Carlo Algorithms

2025-08-24 · Thanh Dang, Mert Gurbuzbalaban, Mohammad Rafiqul Islam, Nian Yao 외 arxiv

Langevin algorithms are popular Markov chain Monte Carlo (MCMC) methods for large-scale sampling problems that often arise in data science. We propose Monte Carlo algorithms based on the discretizations of $P$-th order L…