paper-with-me

Papers

Learning Optimal Admission Control in Partially Observable Queueing Networks

2023-08-04 · Jonatha Anselmi, Bruno Gaujal, Louis-Sébastien Rebuffi

We present an efficient reinforcement learning algorithm that learns the optimal admission control policy in a partially observable queueing network. Specifically, only the arrival and departure times from the network are observable, and optimality refers to the average holding/rejection cost in infinite horizon. While reinforcement learning in Partially Observable Markov Decision Processes (POMDP) is prohibitively expensive in general, we show that our algorithm has a regret that only depends sub-linearly on the maximal number of jobs in the network, $S$. In particular, in contrast with existing regret analyses, our regret bound does not depend on the diameter of the underlying Markov Decision Process (MDP), which in most queueing systems is at least exponential in $S$. The novelty of our approach is to leverage Norton's equivalent theorem for closed product-form queueing networks and an efficient reinforcement learning algorithm for MDPs with the structure of birth-and-death processes.

📄 PDF Abstract BibTeX arXiv:2308.02391

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement Learning

Similar Papers 제목 키워드 기반

Approximate Control for Continuous-Time POMDPs

2024-02-02 · Yannick Eich, Bastian Alt, Heinz Koeppl

This work proposes a decision-making framework for partially observable systems in continuous time with discrete state and action spaces. As optimal decision-making becomes intractable for large state spaces we employ ap…

Decision Making

Arrival Control in Quasi-Reversible Queueing Systems: Optimization and Reinforcement Learning

2025-05-22 · Céline Comte, Pascal Moyal

In this paper, we introduce a versatile scheme for optimizing the arrival rates of quasi-reversible queueing systems. We first propose an alternative definition of quasi-reversibility that encompasses reversibility and h…

Dynamic Control of Service Systems with Returns: Application to Design of Post-Discharge Hospital Readmission Prevention Programs

2022-02-28 · Timothy C. Y. Chan, Simon Y. Huang, Vahid Sarhangian

We study a control problem for queueing systems where customers may return for additional episodes of service after their initial service completion. At each service completion epoch, the decision maker can choose to red…

Probabilistic inverse optimal control for non-linear partially observable systems disentangles perceptual uncertainty and behavioral costs

2023-03-29 · NeurIPS 2023 11 · Dominik Straub, Matthias Schultheis, Heinz Koeppl, Constantin A. Rothkopf

Inverse optimal control can be used to characterize behavior in sequential decision-making tasks. Most existing work, however, is limited to fully observable or linear systems, or requires the action signals to be known.…

Active LearningDecision MakingDecision Making Under UncertaintyImitation Learning+1

Optimal Control of Logically Constrained Partially Observable and Multi-Agent Markov Decision Processes

2023-05-24 · Krishna C. Kalagarla, Dhruva Kartik, Dongming Shen, Rahul Jain 외

Autonomous systems often have logical constraints arising, for example, from safety, operational, or regulatory requirements. Such constraints can be expressed using temporal logic specifications. The system state is oft…