paper-with-me

Papers

Robust Learning for Smoothed Online Convex Optimization with Feedback Delay

2023-10-31 · NeurIPS 2023 11

We study a challenging form of Smoothed Online Convex Optimization, a.k.a. SOCO, including multi-step nonlinear switching costs and feedback delay. We propose a novel machine learning (ML) augmented online algorithm, Robustness-Constrained Learning (RCL), which combines untrusted ML predictions with a trusted expert online algorithm via constrained projection to robustify the ML prediction. Specifically,we prove that RCL is able to guarantee$(1+\lambda)$-competitiveness against any given expert for any$\lambda>0$, while also explicitly training the ML model in a robustification-aware manner to improve the average-case performance. Importantly,RCL is the first ML-augmented algorithm with a provable robustness guarantee in the case of multi-step switching cost and feedback delay.We demonstrate the improvement of RCL in both robustness and average performance using battery management for electrifying transportationas a case study.

📄 PDF Abstract BibTeX arXiv:2310.20098

Code (0)

등록된 구현이 없습니다.

Tasks

Management

Similar Papers 제목 키워드 기반

Online Strongly Convex Optimization with Unknown Delays

2021-03-21 · Yuanyu Wan, Wei-Wei Tu, Lijun Zhang

We investigate the problem of online convex optimization with unknown delays, in which the feedback of a decision arrives with an arbitrary delay. Previous studies have presented a delayed variant of online gradient desc…

Capacity-Constrained Online Convex Optimization with Delayed Feedback

2026-06-10 · Alexander Ryabchenko, Idan Attias, Daniel M. Roy arxiv

Online learning with delayed feedback typically assumes that the learner can track all pending rounds until their feedback arrives. In practice, tracking resources are finite, and feedback from untracked rounds is perman…

A Reduction from Delayed to Immediate Feedback for Online Convex Optimization with Improved Guarantees

2026-02-02 · Alexander Ryabchenko, Idan Attias, Daniel M. Roy arxiv

We develop a reduction-based framework for online learning with delayed feedback that recovers and improves upon existing results for both first-order and bandit convex optimization. Our approach introduces a continuous-…

Revisiting Multi-Agent Asynchronous Online Optimization with Delays: the Strongly Convex Case

2025-03-13 · Lingchan Bao, Tong Wei, Yuanyu Wan

We revisit multi-agent asynchronous online optimization with delays, where only one of the agents becomes active for making the decision at each round, and the corresponding feedback is received by all the agents after u…

Online Sequential Decision-Making with Unknown Delays

2024-02-12 · Ping Wu, Heyan Huang, Zhengyang Liu

In the field of online sequential decision-making, we address the problem with delays utilizing the framework of online convex optimization (OCO), where the feedback of a decision can arrive with an unknown delay. Unlike…

Decision MakingSequential Decision Making