paper-with-me

Papers

A Finite-State Controller Based Offline Solver for Deterministic POMDPs

2025-05-01 · Alex Schutz, Yang You, Matias Mattamala, Ipek Caliskanelli, Bruno Lacerda, Nick Hawes

Deterministic partially observable Markov decision processes (DetPOMDPs) often arise in planning problems where the agent is uncertain about its environmental state but can act and observe deterministically. In this paper, we propose DetMCVI, an adaptation of the Monte Carlo Value Iteration (MCVI) algorithm for DetPOMDPs, which builds policies in the form of finite-state controllers (FSCs). DetMCVI solves large problems with a high success rate, outperforming existing baselines for DetPOMDPs. We also verify the performance of the algorithm in a real-world mobile robot forest mapping scenario.

📄 PDF Abstract BibTeX arXiv:2505.00596

Code (1)

ori-goals/DetMCVI 공식 구현

Similar Papers 제목 키워드 기반

Monte-Carlo Search for an Equilibrium in Dec-POMDPs

2023-05-19 · Yang You, Vincent Thomas, Francis Colas, Olivier Buffet

Decentralized partially observable Markov decision processes (Dec-POMDPs) formalize the problem of designing individual controllers for a group of collaborative agents under stochastic dynamics and partial observability.…

Finite-State Controllers for (Hidden-Model) POMDPs using Deep Reinforcement Learning

2026-02-09 · David Hudák, Maris F. L. Galesloot, Martin Tappler, Martin Kurečka 외 arxiv

Solving partially observable Markov decision processes (POMDPs) requires computing policies under imperfect state information. Despite recent advances, the scalability of existing POMDP solvers remains limited. Moreover,…

Reinforcement Learning

Competitive Control

2021-07-28 · Gautam Goel, Babak Hassibi

We consider control from the perspective of competitive analysis. Unlike much prior work on learning-based control, which focuses on minimizing regret against the best controller selected in hindsight from some specific …

Model Predictive Control

Periodic Finite State Controllers for Efficient POMDP and DEC-POMDP Planning

2011-12-01 · NeurIPS 2011 12 · Joni K. Pajarinen, Jaakko Peltonen

Applications such as robot control and wireless communication require planning under uncertainty. Partially observable Markov decision processes (POMDPs) plan policies for single agents under uncertainty and their decent…

Synergistic Offline-Online Control Synthesis via Local Gaussian Process Regression

2021-10-11 · John Jackson, Luca Laurenti, Eric Frew, Morteza Lahijanian

Autonomous systems often have complex and possibly unknown dynamics due to, e.g., black-box components. This leads to unpredictable behaviors and makes control design with performance guarantees a major challenge. This p…

regression