paper-with-me

Papers

Online Sequential Decision-Making with Unknown Delays

2024-02-12 · Ping Wu, Heyan Huang, Zhengyang Liu

In the field of online sequential decision-making, we address the problem with delays utilizing the framework of online convex optimization (OCO), where the feedback of a decision can arrive with an unknown delay. Unlike previous research that is limited to Euclidean norm and gradient information, we propose three families of delayed algorithms based on approximate solutions to handle different types of received feedback. Our proposed algorithms are versatile and applicable to universal norms. Specifically, we introduce a family of Follow the Delayed Regularized Leader algorithms for feedback with full information on the loss function, a family of Delayed Mirror Descent algorithms for feedback with gradient information on the loss function and a family of Simplified Delayed Mirror Descent algorithms for feedback with the value information of the loss function's gradients at corresponding decision points. For each type of algorithm, we provide corresponding regret bounds under cases of general convexity and relative strong convexity, respectively. We also demonstrate the efficiency of each algorithm under different norms through concrete examples. Furthermore, our theoretical results are consistent with the current best bounds when degenerated to standard settings.

📄 PDF Abstract BibTeX arXiv:2402.07703

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingSequential Decision Making

Similar Papers 제목 키워드 기반

Distributed Online Submodular Maximization under Communication Delays: A Simultaneous Decision-Making Approach

2026-03-29 · Zirui Xu, Vasileios Tzoumas arxiv

We provide a distributed online algorithm for multi-agent submodular maximization under communication delays. We are motivated by the future distributed information-gathering tasks in unknown and dynamic environments, wh…

A Reduction-based Framework for Sequential Decision Making with Delayed Feedback

2023-02-03 · NeurIPS 2023 11 · Yunchang Yang, Han Zhong, Tianhao Wu, Bin Liu 외

We study stochastic delayed feedback in general multi-agent sequential decision making, which includes bandits, single-agent Markov decision processes (MDPs), and Markov games (MGs). We propose a novel reduction-based fr…

Decision MakingSequential Decision Making

Stochastic Online Shortest Path Routing: The Value of Feedback

2013-09-27 · M. Sadegh Talebi, Zhenhua Zou, Richard Combes, Alexandre Proutiere 외

This paper studies online shortest path routing over multi-hop networks. Link costs or delays are time-varying and modeled by independent and identically distributed random processes, whose parameters are initially unkno…

Sequential Decision Making with Expert Demonstrations under Unobserved Heterogeneity

2024-04-10 · Vahid Balazadeh, Keertana Chidambaram, Viet Nguyen, Rahul G. Krishnan 외

We study the problem of online sequential decision-making given auxiliary demonstrations from experts who made their decisions based on unobserved contextual information. These demonstrations can be viewed as solving rel…

Decision MakingMeta Reinforcement LearningMulti-Armed Banditsreinforcement-learning+3

Stochastic Sequential Decision Making over Expanding Networks with Graph Filtering

2026-03-19 · Zhan Gao, Bishwadeep Das, Elvin Isufi arxiv

Graph filters leverage topological information to process networked data with existing methods mainly studying fixed graphs, ignoring that graphs often expand as nodes continually attach with an unknown pattern. The latt…

Multi-agent Reinforcement LearningGraph Neural NetworkDecision Making