paper-with-me

홈 › Papers

Stochastic Approximation with Block Coordinate Optimal Stepsizes

2025-07-11 · Tao Jiang, Lin Xiao arxiv

We consider stochastic approximation with block-coordinate stepsizes and propose adaptive stepsize rules that aim to minimize the expected distance from the next iterate to an (unknown) target point. These stepsize rules employ online estimates of the second moment of the search direction along each block coordinate. The popular Adam algorithm can be interpreted as a variant with a specific estimator. By leveraging a simple conditional estimator, we derive a new method that obtains competitive performance against Adam but requires less memory and fewer hyper-parameters. We prove that this family of methods converges almost surely to a small neighborhood of the target point, and the radius of the neighborhood depends on the bias and variance of the second-moment estimator. Our analysis relies on a simple aiming condition that assumes neither convexity nor smoothness, thus has broad applicability.

📄 PDF Abstract BibTeX arXiv:2507.08963

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Doubly Random Parallel Stochastic Methods for Large Scale Learning

2016-03-22 · Aryan Mokhtari, Alec Koppel, Alejandro Ribeiro

We consider learning problems over training sets in which both, the number of training examples and the dimension of the feature vectors, are large. To solve these problems we propose the random parallel stochastic algor…

Whittle Index Learning Algorithms for Restless Bandits with Constant Stepsizes

2024-09-06 · Vishesh Mittal, Rahul Meshram, Surya Prakash

We study the Whittle index learning algorithm for restless multi-armed bandits. We consider index learning algorithm with Q-learning. We first present Q-learning algorithm with exploration policies -- epsilon-greedy, sof…

Multi-Armed BanditsQ-Learning

Non-Asymptotic Guarantees for Average-Reward Q-Learning with Adaptive Stepsizes

2025-04-25 · Zaiwei Chen

This work presents the first finite-time analysis for the last-iterate convergence of average-reward Q-learning with an asynchronous implementation. A key feature of the algorithm we study is the use of adaptive stepsize…

Q-Learning

Accelerated, Parallel and Proximal Coordinate Descent

2013-12-20 · Olivier Fercoq, Peter Richtárik

We propose a new stochastic coordinate descent method for minimizing the sum of convex functions each of which depends on a small number of coordinates only. Our method (APPROX) is simultaneously Accelerated, Parallel an…

Two-Timescale Linear Stochastic Approximation: Constant Stepsizes Go a Long Way

2024-10-16 · Jeongyeol Kwon, Luke Dotson, Yudong Chen, Qiaomin Xie

Previous studies on two-timescale stochastic approximation (SA) mainly focused on bounding mean-squared errors under diminishing stepsize schemes. In this work, we investigate {\it constant} stpesize schemes through the …