paper-with-me

홈 › Papers

Composing Diverse Policies for Temporally Extended Tasks

2019-07-18 · Daniel Angelov, Yordan Hristov, Michael Burke, Subramanian Ramamoorthy

Robot control policies for temporally extended and sequenced tasks are often characterized by discontinuous switches between different local dynamics. These change-points are often exploited in hierarchical motion planning to build approximate models and to facilitate the design of local, region-specific controllers. However, it becomes combinatorially challenging to implement such a pipeline for complex temporally extended tasks, especially when the sub-controllers work on different information streams, time scales and action spaces. In this paper, we introduce a method that can compose diverse policies comprising motion planning trajectories, dynamic motion primitives and neural network controllers. We introduce a global goal scoring estimator that uses local, per-motion primitive dynamics models and corresponding activation state-space sets to sequence diverse policies in a locally optimal fashion. We use expert demonstrations to convert what is typically viewed as a gradient-based learning process into a planning process without explicitly specifying pre- and post-conditions. We first illustrate the proposed framework using an MDP benchmark to showcase robustness to action and model dynamics mismatch, and then with a particularly complex physical gear assembly task, solved on a PR2 robot. We show that the proposed approach successfully discovers the optimal sequence of controllers and solves both tasks efficiently.

📄 PDF Abstract BibTeX arXiv:1907.08199

Code (0)

등록된 구현이 없습니다.

Tasks

Hierarchical Reinforcement LearningMotion Planning

Similar Papers 제목 키워드 기반

Planning with Goal-Conditioned Policies

2019-11-19 · NeurIPS 2019 12 · Soroush Nasiriany, Vitchyr H. Pong, Steven Lin, Sergey Levine

Planning methods can solve temporally extended sequential decision making problems by composing simple behaviors. However, planning requires suitable abstractions for the states and transitions, which typically need to b…

Decision Makingreinforcement-learningReinforcement LearningReinforcement Learning (RL)+3

Unveiling Options with Neural Decomposition

2024-10-15 · Mahdi Alikhasi, Levi H. S. Lelis

In reinforcement learning, agents often learn policies for specific tasks without the ability to generalize this knowledge to related tasks. This paper introduces an algorithm that attempts to address this limitation by …

Temporally Extended Successor Representations

2022-09-25 · Matthew J. Sargent, Peter J. Bentley, Caswell Barry, William de Cothi

We present a temporally extended variation of the successor representation, which we term t-SR. t-SR captures the expected state transition dynamics of temporally extended actions by constructing successor representation…

Planning to Practice: Efficient Online Fine-Tuning by Composing Goals in Latent Space

2022-05-17 · Kuan Fang, Patrick Yin, Ashvin Nair, Sergey Levine

General-purpose robots require diverse repertoires of behaviors to complete challenging tasks in real-world unstructured environments. To address this issue, goal-conditioned reinforcement learning aims to acquire polici…

reinforcement-learningReinforcement Learning (RL)

Active Hierarchical Exploration with Stable Subgoal Representation Learning

2021-05-31 · ICLR 2022 4 · Siyuan Li, Jin Zhang, Jianhao Wang, Yang Yu 외

Goal-conditioned hierarchical reinforcement learning (GCHRL) provides a promising approach to solving long-horizon tasks. Recently, its success has been extended to more general settings by concurrently learning hierarch…

continuous-controlContinuous ControlHierarchical Reinforcement LearningRepresentation Learning