paper-with-me

Papers

Compactly Restrictable Metric Policy Optimization Problems

2022-07-12 · Victor D. Dorobantu, Kamyar Azizzadenesheli, Yisong Yue

We study policy optimization problems for deterministic Markov decision processes (MDPs) with metric state and action spaces, which we refer to as Metric Policy Optimization Problems (MPOPs). Our goal is to establish theoretical results on the well-posedness of MPOPs that can characterize practically relevant continuous control systems. To do so, we define a special class of MPOPs called Compactly Restrictable MPOPs (CR-MPOPs), which are flexible enough to capture the complex behavior of robotic systems but specific enough to admit solutions using dynamic programming methods such as value iteration. We show how to arrive at CR-MPOPs using forward-invariance. We further show that our theoretical results on CR-MPOPs can be used to characterize feedback linearizable control affine systems.

📄 PDF Abstract BibTeX arXiv:2207.05850

Code (0)

등록된 구현이 없습니다.

Tasks

continuous-controlContinuous Control

Similar Papers 제목 키워드 기반

Generative Modelling of Stochastic Actions with Arbitrary Constraints in Reinforcement Learning

2023-11-26 · NeurIPS 2023 11 · Changyu Chen, Ramesha Karunasena, Thanh Hong Nguyen, Arunesh Sinha 외

Many problems in Reinforcement Learning (RL) seek an optimal policy with large discrete multidimensional yet unordered action spaces; these include problems in randomized allocation of resources such as placements of mul…

reinforcement-learningReinforcement Learning (RL)valid

Efficient Global Planning in Large MDPs via Stochastic Primal-Dual Optimization

2022-10-21 · Gergely Neu, Nneka Okolo

We propose a new stochastic primal-dual optimization algorithm for planning in a large discounted Markov decision process with a generative model and linear function approximation. Assuming that the feature map approxima…

Solving Transition-Independent Multi-agent MDPs with Sparse Interactions (Extended version)

2015-11-29 · Joris Scharpff, Diederik M. Roijers, Frans A. Oliehoek, Matthijs T. J. Spaan 외

In cooperative multi-agent sequential decision making under uncertainty, agents must coordinate to find an optimal joint policy that maximises joint value. Typical algorithms exploit additive structure in the value funct…

Decision MakingDecision Making Under UncertaintySequential Decision Making

Sparse Gaussian Processes via Parametric Families of Compactly-supported Kernels

2020-06-05 · Jarred Barber

Gaussian processes are powerful models for probabilistic machine learning, but are limited in application by their $O(N^3)$ inference complexity. We propose a method for deriving parametric families of kernel functions w…

Gaussian Processes

Optimized Product Quantization for Approximate Nearest Neighbor Search

2013-06-01 · CVPR 2013 6 · Tiezheng Ge, Kaiming He, Qifa Ke, Jian Sun

Product quantization is an effective vector quantization approach to compactly encode high-dimensional vectors for fast approximate nearest neighbor (ANN) search. The essence of product quantization is to decompose the o…

Quantization