paper-with-me

Papers

Geometry-Aware Universal Mirror-Prox

2020-11-23 · Reza Babanezhad, Simon Lacoste-Julien

Mirror-prox (MP) is a well-known algorithm to solve variational inequality (VI) problems. VI with a monotone operator covers a large group of settings such as convex minimization, min-max or saddle point problems. To get a convergent algorithm, the step-size of the classic MP algorithm relies heavily on the problem dependent knowledge of the operator such as its smoothness parameter which is hard to estimate. Recently, a universal variant of MP for smooth/bounded operators has been introduced that depends only on the norm of updates in MP. In this work, we relax the dependence to evaluating the norm of updates to Bregman divergence between updates. This relaxation allows us to extends the analysis of universal MP to the settings where the operator is not smooth or bounded. Furthermore, we analyse the VI problem with a stochastic monotone operator in different settings and obtain an optimal rate up to a logarithmic factor.

📄 PDF Abstract BibTeX arXiv:2011.11203

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

A Universal Banach--Bregman Framework for Stochastic Iterations: Unifying Stochastic Mirror Descent, Learning and LLM Training

2025-09-17 · Johnny R. Zhang, Xiaomei Mi, Gaoyuan Du, Qianyi Sun 외 arxiv

Stochastic optimization powers the scalability of modern artificial intelligence, spanning machine learning, deep learning, reinforcement learning, and large language model training. Yet, existing theory remains largely …

Stochastic OptimizationReinforcement LearningSparse Learning

Mirror-Fusion Attention for Reflection-Aware Self-Supervised Representation Learning

2026-07-01 · Ruixin Li, Jin Liu, Yuling Shi, Stefano Lodi arxiv

Most self-supervised learning (SSL) methods encourage invariance across augmentations, but strict flip invariance can suppress informative left--right correspondences in approximately bilateral data such as medical image…

Self-Supervised LearningRepresentation Learning

Behavior-Induced Mirror-Prox Temporal-Difference Learning for Faster Off-Policy Prediction

2026-05-16 · Xingguo Chen, Yuchen Shen, Shangdong Yang, Chao Li 외 arxiv

Gradient temporal-difference methods provide stable off-policy prediction with linear function approximation, but their practical performance is strongly affected by the geometry induced by the auxiliary-variable metric.…

Variational Online Mirror Descent for Robust Learning in Schrödinger Bridge

2025-04-03 · Dong-Sig Han, Jaein Kim, Hee Bin Yoo, Byoung-Tak Zhang

Sch\"odinger bridge (SB) has evolved into a universal class of probabilistic generative models. In practice, however, estimated learning signals are often uncertain, and the reliability promised by existing methods is of…

The Information Geometry of Mirror Descent

2013-10-29 · Garvesh Raskutti, Sayan Mukherjee

Information geometry applies concepts in differential geometry to probability and statistics and is especially useful for parameter estimation in exponential families where parameters are known to lie on a Riemannian man…

parameter estimation