paper-with-me

홈 › Papers

Opportunistic Qualitative Planning in Stochastic Systems with Incomplete Preferences over Reachability Objectives

2022-10-04 · Abhishek N. Kulkarni, Jie Fu

Preferences play a key role in determining what goals/constraints to satisfy when not all constraints can be satisfied simultaneously. In this paper, we study how to synthesize preference satisfying plans in stochastic systems, modeled as an MDP, given a (possibly incomplete) combinative preference model over temporally extended goals. We start by introducing new semantics to interpret preferences over infinite plays of the stochastic system. Then, we introduce a new notion of improvement to enable comparison between two prefixes of an infinite play. Based on this, we define two solution concepts called safe and positively improving (SPI) and safe and almost-surely improving (SASI) that enforce improvements with a positive probability and with probability one, respectively. We construct a model called an improvement MDP, in which the synthesis of SPI and SASI strategies that guarantee at least one improvement reduces to computing positive and almost-sure winning strategies in an MDP. We present an algorithm to synthesize the SPI and SASI strategies that induce multiple sequential improvements. We demonstrate the proposed approach using a robot motion planning problem.

📄 PDF Abstract BibTeX arXiv:2210.01878

Code (0)

등록된 구현이 없습니다.

Tasks

Motion Planning

Similar Papers 제목 키워드 기반

Iterative Motion Planning in Multi-agent Systems with Opportunistic Communication under Disturbance

2025-03-16 · Neelanga Thelasingha, Agung Julius, James Humann, James Dotterweich

In complex multi-agent systems involving heterogeneous teams, uncertainty arises from numerous sources like environmental disturbances, model inaccuracies, and changing tasks. This causes planned trajectories to become i…

Motion Planning

Probabilistic Planning with Preferences over Temporal Goals

2021-03-26 · Jie Fu

We present a formal language for specifying qualitative preferences over temporal goals and a preference-based planning method in stochastic systems. Using automata-theoretic modeling, the proposed specification allows u…

Temporal Sequences

Reinforcement Learning with Budget-Constrained Nonparametric Function Approximation for Opportunistic Spectrum Access

2017-06-14 · Theodoros Tsiligkaridis, David Romero

Opportunistic spectrum access is one of the emerging techniques for maximizing throughput in congested bands and is enabled by predicting idle slots in spectrum. We propose a kernel-based reinforcement learning approach …

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Cloth Manipulation Planning on Basis of Mesh Representations with Incomplete Domain Knowledge and Voxel-to-Mesh Estimation

2021-03-15 · Solvi Arnold, Daisuke Tanaka, Kimitoshi Yamazaki

We consider the problem of open-goal planning for robotic cloth manipulation. Core of our system is a neural network trained as a forward model of cloth behaviour under manipulation, with planning performed through backp…

Reinforcement Learning for Opportunistic Routing in Software-Defined LEO-Terrestrial Systems

2026-01-20 · Sivaram Krishnan, Zhouyou Gu, Jihong Park, Sung-Min Oh 외 arxiv

The proliferation of large-scale low Earth orbit (LEO) satellite constellations is driving the need for intelligent routing strategies that can effectively deliver data to terrestrial networks under rapidly time-varying …

Stochastic OptimizationReinforcement Learning