paper-with-me

홈 › Papers

A weak convergence approach to large deviations for stochastic approximations

2025-02-04 · Henrik Hult, Adam Lindhe, Pierre Nyquist, Guo-Jhen Wu

The theory of stochastic approximations form the theoretical foundation for studying convergence properties of many popular recursive learning algorithms in statistics, machine learning and statistical physics. Large deviations for stochastic approximations provide asymptotic estimates of the probability that the learning algorithm deviates from its expected path, given by a limit ODE, and the large deviation rate function gives insights to the most likely way that such deviations occur. In this paper we prove a large deviation principle for general stochastic approximations with state-dependent Markovian noise and decreasing step size. Using the weak convergence approach to large deviations, we generalize previous results for stochastic approximations and identify the appropriate scaling sequence for the large deviation principle. We also give a new representation for the rate function, in which the rate function is expressed as an action functional involving the family of Markov transition kernels. Examples of learning algorithms that are covered by the large deviation principle include stochastic gradient descent, persistent contrastive divergence and the Wang-Landau algorithm.

📄 PDF Abstract BibTeX arXiv:2502.02529

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Q-Learning for Stochastic Control under General Information Structures and Non-Markovian Environments

2023-10-31 · Ali Devran Kara, Serdar Yuksel

As a primary contribution, we present a convergence theorem for stochastic iterations, and in particular, Q-learning iterates, under a general, possibly non-Markovian, stochastic environment. Our conditions for convergen…

Q-LearningQuantization

Weak error rates of numerical schemes for rough volatility

2022-03-17 · Paul Gassiat

Simulation of rough volatility models involves discretization of stochastic integrals where the integrand is a function of a (correlated) fractional Brownian motion of Hurst index $H \in (0,1/2)$. We obtain results on th…

Diffusion Approximations for Thompson Sampling

2021-05-19 · Lin Fan, Peter W. Glynn

We study the behavior of Thompson sampling from the perspective of weak convergence. In the regime with small $\gamma > 0$, where the gaps between arm means scale as $\sqrt{\gamma}$ and over time horizons that scale as $…

Multi-Armed BanditsThompson Sampling

Subgradient Selection Convergence Implies Uniform Subdifferential Set Convergence: And Other Tight Convergences Rates in Stochastic Convex Composite Minimization

2024-05-16 · Feng Ruan

In nonsmooth, nonconvex stochastic optimization, understanding the uniform convergence of subdifferential mappings is crucial for analyzing stationary points of sample average approximations of risk as they approach the …

Stochastic Optimization

Weak Convergence Properties of Constrained Emphatic Temporal-difference Learning with Constant and Slowly Diminishing Stepsize

2015-11-23 · Huizhen Yu

We consider the emphatic temporal-difference (TD) algorithm, ETD($\lambda$), for learning the value functions of stationary policies in a discounted, finite state and action Markov decision process. The ETD($\lambda$) al…