paper-with-me

홈 › Papers

Expectation Optimization with Probabilistic Guarantees in POMDPs with Discounted-sum Objectives

2018-04-27 · Krishnendu Chatterjee, Adrián Elgyütt, Petr Novotný, Owen Rouillé

Partially-observable Markov decision processes (POMDPs) with discounted-sum payoff are a standard framework to model a wide range of problems related to decision making under uncertainty. Traditionally, the goal has been to obtain policies that optimize the expectation of the discounted-sum payoff. A key drawback of the expectation measure is that even low probability events with extreme payoff can significantly affect the expectation, and thus the obtained policies are not necessarily risk-averse. An alternate approach is to optimize the probability that the payoff is above a certain threshold, which allows obtaining risk-averse policies, but ignores optimization of the expectation. We consider the expectation optimization with probabilistic guarantee (EOPG) problem, where the goal is to optimize the expectation ensuring that the payoff is above a given threshold with at least a specified probability. We present several results on the EOPG problem, including the first algorithm to solve it.

📄 PDF Abstract BibTeX arXiv:1804.10601

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingDecision Making Under Uncertainty

Similar Papers 제목 키워드 기반

Optimizing Expectation with Guarantees in POMDPs (Technical Report)

2016-11-26 · Krishnendu Chatterjee, Petr Novotný, Guillermo A. Pérez, Jean-François Raskin 외

A standard objective in partially-observable Markov decision processes (POMDPs) is to find a policy that maximizes the expected discounted-sum payoff. However, such policies may still permit unlikely but highly undesirab…

Sound Heuristic Search Value Iteration for Undiscounted POMDPs with Reachability Objectives

2024-06-05 · Qi Heng Ho, Martin S. Feather, Federico Rossi, Zachary N. Sunberg 외

Partially Observable Markov Decision Processes (POMDPs) are powerful models for sequential decision making under transition and observation uncertainties. This paper studies the challenging yet important problem in POMDP…

Decision MakingEfficient ExplorationHeuristic SearchSequential Decision Making

The Geometry of Memoryless Stochastic Policy Optimization in Infinite-Horizon POMDPs

2021-10-14 · ICLR 2022 4 · Johannes Müller, Guido Montúfar

We consider the problem of finding the best memoryless stochastic policy for an infinite-horizon partially observable Markov decision process (POMDP) with finite state and action spaces with respect to either the discoun…

Sublinear Regret for Learning POMDPs

2021-07-08 · Yi Xiong, Ningyuan Chen, Xuefeng Gao, Xiang Zhou

We study the model-based undiscounted reinforcement learning for partially observable Markov decision processes (POMDPs). The oracle we consider is the optimal policy of the POMDP with a known environment in terms of the…

reinforcement-learningReinforcement Learning (RL)

Risk-Averse Decision Making Under Uncertainty

2021-09-09 · Mohamadreza Ahmadi, Ugo Rosolia, Michel D. Ingham, Richard M. Murray 외

A large class of decision making under uncertainty problems can be described via Markov decision processes (MDPs) or partially observable MDPs (POMDPs), with application to artificial intelligence and operations research…

Decision MakingDecision Making Under Uncertainty