paper-with-me

Papers

Parameter-free Online Linear Optimization with Side Information via Universal Coin Betting

2022-02-04 · J. Jon Ryu, Alankrita Bhatt, Young-Han Kim

A class of parameter-free online linear optimization algorithms is proposed that harnesses the structure of an adversarial sequence by adapting to some side information. These algorithms combine the reduction technique of Orabona and P{\'a}l (2016) for adapting coin betting algorithms for online linear optimization with universal compression techniques in information theory for incorporating sequential side information to coin betting. Concrete examples are studied in which the side information has a tree structure and consists of quantized values of the previous symbols of the adversarial sequence, including fixed-order and variable-order Markov cases. By modifying the context-tree weighting technique of Willems, Shtarkov, and Tjalkens (1995), the proposed algorithm is further refined to achieve the best performance over all adaptive algorithms with tree-structured side information of a given maximum order in a computationally efficient manner.

📄 PDF Abstract BibTeX arXiv:2202.02406

Code (1)

jongharyu/olo-with-side-information 공식 구현

Similar Papers 제목 키워드 기반

Coin Betting and Parameter-Free Online Learning

2016-02-12 · NeurIPS 2016 12 · Francesco Orabona, Dávid Pál

In the recent years, a number of parameter-free algorithms have been developed for online linear optimization over Hilbert spaces and for learning with expert advice. These algorithms achieve optimal regret bounds that d…

Feedback Linearization for Unknown Systems via Reinforcement Learning

2019-10-29 · Tyler Westenbroek, David Fridovich-Keil, Eric Mazumdar, Shreyas Arora 외

We present a novel approach to control design for nonlinear systems which leverages model-free policy optimization techniques to learn a linearizing controller for a physical plant with unknown dynamics. Feedback lineari…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Parameter-free Mirror Descent

2022-02-26 · Andrew Jacobsen, Ashok Cutkosky

We develop a modified online mirror descent framework that is suitable for building adaptive and parameter-free algorithms in unbounded domains. We leverage this technique to develop the first unconstrained online linear…

Implicit Parameter-free Online Learning with Truncated Linear Models

2022-03-19 · Keyi Chen, Ashok Cutkosky, Francesco Orabona

Parameter-free algorithms are online learning algorithms that do not require setting learning rates. They achieve optimal regret with respect to the distance between the initial point and any competitor. Yet, parameter-f…

Stochastic Optimization

Decentralized Parameter-Free Online Learning

2025-10-17 · Tomas Ortega, Hamid Jafarkhani arxiv

We propose the first parameter-free decentralized online learning algorithms with network regret guarantees, which achieve sublinear regret without requiring hyperparameter tuning. This family of algorithms connects mult…