paper-with-me

홈 › Papers

Point-Based Value Iteration and Approximately Optimal Dynamic Sensor Selection for Linear-Gaussian Processes

2020-12-23 · Michael Hibbard, Kirsten Tuggle, Takashi Tanaka

The problem of synthesizing an optimal sensor selection policy is pertinent to a variety of engineering applications ranging from event detection to autonomous navigation. We consider such a synthesis problem over an infinite time horizon with a discounted cost criterion. We formulate this problem in terms of a value iteration over the continuous space of covariance matrices. To obtain a computationally tractable solution, we subsequently formulate an approximate sensor selection problem, which is solvable through a point-based value iteration over a finite "mesh" of covariance matrices with a user-defined bounded trace. We provide theoretical guarantees bounding the suboptimality of the sensor selection policies synthesized through this method and provide numerical examples comparing them to known results.

📄 PDF Abstract BibTeX arXiv:2012.12842

Code (0)

등록된 구현이 없습니다.

Tasks

Autonomous NavigationEvent DetectionGaussian Processes

Similar Papers 제목 키워드 기반

Approximate Midpoint Policy Iteration for Linear Quadratic Control

2020-11-28 · Benjamin Gravell, Iman Shames, Tyler Summers

We present a midpoint policy iteration algorithm to solve linear quadratic optimal control problems in both model-based and model-free settings. The algorithm is a variation of Newton's method, and we show that in the mo…

A Contracting Dynamical System Perspective toward Interval Markov Decision Processes

2023-09-17 · Saber Jafarpour, Samuel Coogan

Interval Markov decision processes are a class of Markov models where the transition probabilities between the states belong to intervals. In this paper, we study the problem of efficient estimation of the optimal polici…

Preemptive Scheduling of EV Charging for Providing Demand Response Services

2022-08-21 · Shiping Shao, Farshad Harirchi, Devang Dave, Abhishek Gupta

We develop a new algorithm for scheduling the charging process of a large number of electric vehicles (EVs) over a finite horizon. We assume that EVs arrive at the charging stations with different charge levels and diffe…

Scheduling

Forward-PECVaR Algorithm: Exact Evaluation for CVaR SSPs

2023-03-01 · Willy Arthur Silva Reis, Denis Benevolo Pais, Valdinei Freire, Karina Valdivia Delgado

The Stochastic Shortest Path (SSP) problem models probabilistic sequential-decision problems where an agent must pursue a goal while minimizing a cost function. Because of the probabilistic dynamics, it is desired to hav…

Generalized Second Order Value Iteration in Markov Decision Processes

2019-05-10 · Chandramouli Kamanchi, Raghuram Bharadwaj Diddigi, Shalabh Bhatnagar

Value iteration is a fixed point iteration technique utilized to obtain the optimal value function and policy in a discounted reward Markov Decision Process (MDP). Here, a contraction operator is constructed and applied …

Reinforcement Learning