paper-with-me

Papers

Active Trajectory Estimation for Partially Observed Markov Decision Processes via Conditional Entropy

2021-04-04 · Timothy L. Molloy, Girish N. Nair

In this paper, we consider the problem of controlling a partially observed Markov decision process (POMDP) in order to actively estimate its state trajectory over a fixed horizon with minimal uncertainty. We pose a novel active smoothing problem in which the objective is to directly minimise the smoother entropy, that is, the conditional entropy of the (joint) state trajectory distribution of concern in fixed-interval Bayesian smoothing. Our formulation contrasts with prior active approaches that minimise the sum of conditional entropies of the (marginal) state estimates provided by Bayesian filters. By establishing a novel form of the smoother entropy in terms of the POMDP belief (or information) state, we show that our active smoothing problem can be reformulated as a (fully observed) Markov decision process with a value function that is concave in the belief state. The concavity of the value function is of particular importance since it enables the approximate solution of our active smoothing problem using piecewise-linear function approximations in conjunction with standard POMDP solvers. We illustrate the approximate solution of our active smoothing problem in simulation and compare its performance to alternative approaches based on minimising marginal state estimate uncertainties.

📄 PDF Abstract BibTeX arXiv:2104.01545

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Smoother Entropy for Active State Trajectory Estimation and Obfuscation in POMDPs

2021-08-19 · Timothy L. Molloy, Girish N. Nair

We study the problem of controlling a partially observed Markov decision process (POMDP) to either aid or hinder the estimation of its state trajectory. We encode the estimation objectives via the smoother entropy, which…

Entropy-Regularized Partially Observed Markov Decision Processes

2021-12-22 · Timothy L. Molloy, Girish N. Nair

We investigate partially observed Markov decision processes (POMDPs) with cost functions regularized by entropy terms describing state, observation, and control uncertainty. Standard POMDP techniques are shown to offer b…

State Estimation

A finite-sample bound for identifying partially observed linear switched systems from a single trajectory

2025-03-17 · Daniel Racz, Mihaly Petreczky, Balint Daroczy

We derive a finite-sample probabilistic bound on the parameter estimation error of a system identification algorithm for Linear Switched Systems. The algorithm estimates Markov parameters from a single trajectory and app…

parameter estimation

Learning Partially Observed Linear Dynamical Systems from Logarithmic Number of Samples

2020-10-08 · Salar Fattahi

In this work, we study the problem of learning partially observed linear dynamical systems from a single sample trajectory. A major practical challenge in the existing system identification methods is the undesirable dep…

Finite Sample Identification of Partially Observed Bilinear Dynamical Systems

2025-01-13 · Yahya Sattar, Yassir Jedra, Maryam Fazel, Sarah Dean

We consider the problem of learning a realization of a partially observed bilinear dynamical system (BLDS) from noisy input-output data. Given a single trajectory of input-output samples, we provide a finite time analysi…