paper-with-me

Papers

Quantum POMDPs

2014-06-11 · Jennifer Barry, Daniel T. Barry, Scott Aaronson

We present quantum observable Markov decision processes (QOMDPs), the quantum analogues of partially observable Markov decision processes (POMDPs). In a QOMDP, an agent's state is represented as a quantum state and the agent can choose a superoperator to apply. This is similar to the POMDP belief state, which is a probability distribution over world states and evolves via a stochastic matrix. We show that the existence of a policy of at least a certain value has the same complexity for QOMDPs and POMDPs in the polynomial and infinite horizon cases. However, we also prove that the existence of a policy that can reach a goal state is decidable for goal POMDPs and undecidable for goal QOMDPs.

📄 PDF Abstract BibTeX arXiv:1406.2858

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Agent policies from higher-order causal functions

2025-12-11 · Matt Wilson arxiv

We establish a correspondence between equivalence classes of agent-state policies for deterministic POMDPs and one-input process functions (the classical-deterministic limit of higher-order quantum operations). We use th…

Multi-Environment POMDPs: Discrete Model Uncertainty Under Partial Observability

2025-10-27 · Eline M. Bovy, Caleb Probine, Marnix Suilen, Ufuk Topcu 외 arxiv

Multi-environment POMDPs (ME-POMDPs) extend standard POMDPs with discrete model uncertainty. ME-POMDPs represent a finite set of POMDPs that share the same state, action, and observation spaces, but may arbitrarily vary …

QANTIS: A Hardware-Validated Quantum Platform for POMDP Planning and Multi-Target Data Association

2026-02-28 · Bayram Yüksel Eker, Suayb S. Arslan, Özgür Nazlı, Mustafa Serhat Demirgil 외 arxiv

Autonomous navigation under uncertainty requires solving partially observable Markov decision processes (POMDPs) for planning and assigning sensor measurements to tracked targets--a task known as multi-target data associ…

\textsc{rfPG}: Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs

2025-05-14 · Maris F. L. Galesloot, Roman Andriushchenko, Milan Češka, Sebastian Junges 외

Partially observable Markov decision processes (POMDPs) model specific environments in sequential decision-making under uncertainty. Critically, optimal policies for POMDPs may not be robust against perturbations in the …

Decision Making Under UncertaintySequential Decision Making

Finite-State Controllers for (Hidden-Model) POMDPs using Deep Reinforcement Learning

2026-02-09 · David Hudák, Maris F. L. Galesloot, Martin Tappler, Martin Kurečka 외 arxiv

Solving partially observable Markov decision processes (POMDPs) requires computing policies under imperfect state information. Despite recent advances, the scalability of existing POMDP solvers remains limited. Moreover,…

Reinforcement Learning