paper-with-me

Papers

Online Linear Optimization via Smoothing

2014-05-23 · Jacob Abernethy, Chansoo Lee, Abhinav Sinha, Ambuj Tewari

We present a new optimization-theoretic approach to analyzing Follow-the-Leader style algorithms, particularly in the setting where perturbations are used as a tool for regularization. We show that adding a strongly convex penalty function to the decision rule and adding stochastic perturbations to data correspond to deterministic and stochastic smoothing operations, respectively. We establish an equivalence between "Follow the Regularized Leader" and "Follow the Perturbed Leader" up to the smoothness properties. This intuition leads to a new generic analysis framework that recovers and improves the previous known regret bounds of the class of algorithms commonly known as Follow the Perturbed Leader.

📄 PDF Abstract BibTeX arXiv:1405.6076

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Nonlinear Model Predictive Control Based on Constraint-Aware Particle Filtering/Smoothing

2022-05-09 · Iman Askari, Shen Zeng, Huazhen Fang

Nonlinear model predictive control (NMPC) has gained widespread use in many applications. Its formulation traditionally involves repetitively solving a nonlinear constrained optimization problem online. In this paper, we…

Model Predictive Control

Optimization-Induced Graph Implicit Nonlinear Diffusion

2022-06-29 · Qi Chen, Yifei Wang, Yisen Wang, Jiansheng Yang 외

Due to the over-smoothing issue, most existing graph neural networks can only capture limited dependencies with their inherently finite aggregation layers. To overcome this limitation, we propose a new kind of graph conv…

A Recursive Newton Method for Smoothing in Nonlinear State Space Models

2023-06-15 · Fatemeh Yaghoobi, Hany Abdulsamad, Simo Särkkä

In this paper, we use the optimization formulation of nonlinear Kalman filtering and smoothing problems to develop second-order variants of iterated Kalman smoother (IKS) methods. We show that Newton's method corresponds…

State Space Models

Adaptive Matrix Online Learning through Smoothing with Guarantees for Nonsmooth Nonconvex Optimization

2026-02-09 · Ruichen Jiang, Zakaria Mhammedi, Mehryar Mohri, Aryan Mokhtari arxiv

We study online linear optimization with matrix variables constrained by the operator norm, a setting where the geometry renders designing data-dependent and efficient adaptive algorithms challenging. The best-known adap…

Proximal Point Imitation Learning

2022-09-22 · Luca Viano, Angeliki Kamoutsi, Gergely Neu, Igor Krawczuk 외

This work develops new algorithms with rigorous efficiency guarantees for infinite horizon imitation learning (IL) with linear function approximation without restrictive coherence assumptions. We begin with the minimax f…

Imitation Learning