paper-with-me

홈 › Papers

Worst-Case Control and Learning Using Partial Observations Over an Infinite Time-Horizon

2023-03-28 · Aditya Dave, Ioannis Faros, Nishanth Venkatesh, Andreas A. Malikopoulos

Safety-critical cyber-physical systems require control strategies whose worst-case performance is robust against adversarial disturbances and modeling uncertainties. In this paper, we present a framework for approximate control and learning in partially observed systems to minimize the worst-case discounted cost over an infinite time horizon. We model disturbances to the system as finite-valued uncertain variables with unknown probability distributions. For problems with known system dynamics, we construct a dynamic programming (DP) decomposition to compute the optimal control strategy. Our first contribution is to define information states that improve the computational tractability of this DP without loss of optimality. Then, we describe a simplification for a class of problems where the incurred cost is observable at each time instance. Our second contribution is defining an approximate information state that can be constructed or learned directly from observed data for problems with observable costs. We derive bounds on the performance loss of the resulting approximate control strategy and illustrate the effectiveness of our approach in partially observed decision-making problems with a numerical example.

📄 PDF Abstract BibTeX arXiv:2303.16321

Code (0)

등록된 구현이 없습니다.

Tasks

Decision Making

Similar Papers 제목 키워드 기반

Safe Control of Partially-Observed Linear Time-Varying Systems with Minimal Worst-Case Dynamic Regret

2022-08-18 · HongYu Zhou, Vasileios Tzoumas

We present safe control of partially-observed linear time-varying systems in the presence of unknown and unpredictable process and measurement noise. We introduce a control algorithm that minimizes dynamic regret, i.e., …

When Is Partially Observable Reinforcement Learning Not Scary?

2022-04-19 · Qinghua Liu, Alan Chung, Csaba Szepesvári, Chi Jin

Applications of Reinforcement Learning (RL), in which agents learn to make a sequence of decisions despite lacking complete information about the latent states of the controlled system, that is, they act under partial ob…

Partially Observable Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Lifting DecPOMDPs for Nanoscale Systems -- A Work in Progress

2021-10-18 · Tanya Braun, Stefan Fischer, Florian Lau, Ralf Möller

DNA-based nanonetworks have a wide range of promising use cases, especially in the field of medicine. With a large set of agents, a partially observable stochastic environment, and noisy observations, such nanoscale syst…

Regret-Optimal Control under Partial Observability

2023-11-10 · Joudi Hajar, Oron Sabag, Babak Hassibi

This paper studies online solutions for regret-optimal control in partially observable systems over an infinite-horizon. Regret-optimal control aims to minimize the difference in LQR cost between causal and non-causal co…

Complexity of Decentralized Control: Special Cases

2009-12-01 · NeurIPS 2009 12 · Martin Allen, Shlomo Zilberstein

The worst-case complexity of general decentralized POMDPs, which are equivalent to partially observable stochastic games (POSGs) is very high, both for the cooperative and competitive cases. Some reductions in complexit…