paper-with-me

홈 › Papers

Structural Estimation of Partially Observable Markov Decision Processes

2020-08-02 · Yanling Chang, Alfredo Garcia, Zhide Wang, Lu Sun

In many practical settings control decisions must be made under partial/imperfect information about the evolution of a relevant state variable. Partially Observable Markov Decision Processes (POMDPs) is a relatively well-developed framework for modeling and analyzing such problems. In this paper we consider the structural estimation of the primitives of a POMDP model based upon the observable history of the process. We analyze the structural properties of POMDP model with random rewards and specify conditions under which the model is identifiable without knowledge of the state dynamics. We consider a soft policy gradient algorithm to compute a maximum likelihood estimator and provide a finite-time characterization of convergence to a stationary point. We illustrate the estimation methodology with an application to optimal equipment replacement. In this context, replacement decisions must be made under partial/imperfect information on the true state (i.e. condition of the equipment). We use synthetic and real data to highlight the robustness of the proposed methodology and characterize the potential for misspecification when partial state observability is ignored.

📄 PDF Abstract BibTeX arXiv:2008.00500

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Hidden Markov Model Estimation-Based Q-learning for Partially Observable Markov Decision Process

2018-09-17 · Hyung-Jin Yoon, Donghwan Lee, Naira Hovakimyan

The objective is to study an on-line Hidden Markov model (HMM) estimation-based Q-learning algorithm for partially observable Markov decision process (POMDP) on finite state and action sets. When the full state observati…

Q-Learning

Partially Observable RL with B-Stability: Unified Structural Condition and Sharp Sample-Efficient Algorithms

2022-09-29 · Fan Chen, Yu Bai, Song Mei

Partial Observability -- where agents can only observe partial information about the true underlying state of the system -- is ubiquitous in real-world applications of Reinforcement Learning (RL). Theoretically, learning…

Reinforcement Learning (RL)

Decision Making for Autonomous Vehicles

2023-04-27 · Xinchen Li, Levent Guvenc, Bilin Aksun-Guvenc

This paper is on decision making of autonomous vehicles for handling roundabouts. The round intersection is introduced first followed by the Markov Decision Processes (MDP), the Partially Observable Markov Decision Proce…

Autonomous VehiclesDecision Making

End-to-End Policy Gradient Method for POMDPs and Explainable Agents

2023-04-19 · Soichiro Nishimori, Sotetsu Koyamada, Shin Ishii

Real-world decision-making problems are often partially observable, and many can be formulated as a Partially Observable Markov Decision Process (POMDP). When we apply reinforcement learning (RL) algorithms to the POMDP,…

Autonomous DrivingDecision Makingreinforcement-learningReinforcement Learning (RL)

Weathering Ongoing Uncertainty: Learning and Planning in a Time-Varying Partially Observable Environment

2023-12-06 · Gokul Puthumanaillam, Xiangyu Liu, Negar Mehr, Melkior Ornik

Optimal decision-making presents a significant challenge for autonomous systems operating in uncertain, stochastic and time-varying environments. Environmental variability over time can significantly impact the system's …

Decision MakingState Estimation