paper-with-me

Papers

Skill Discovery in Continuous Reinforcement Learning Domains using Skill Chaining

2009-12-01 · NeurIPS 2009 12 · George Konidaris, Andrew G. Barto

We introduce skill chaining, a skill discovery method for reinforcement learning agents in continuous domains, that builds chains of skills leading to an end-of-task reward. We demonstrate experimentally that it creates skills that result in performance benefits in a challenging continuous domain.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

Option Discovery using Deep Skill Chaining

2020-05-01 · ICLR 2020 1 · Akhil Bagaria, George Konidaris

Autonomously discovering temporally extended actions, or skills, is a longstanding goal of hierarchical reinforcement learning. We propose a new algorithm that combines skill chaining with deep neural networks to autonom…

continuous-controlContinuous ControlHierarchical Reinforcement Learningreinforcement-learning+2

Unsupervised Discovery of Continuous Skills on a Sphere

2023-05-21 · Takahisa Imagawa, Takuya Hiraoka, Yoshimasa Tsuruoka

Recently, methods for learning diverse skills to generate various behaviors without external rewards have been actively studied as a form of unsupervised reinforcement learning. However, most of the existing methods lear…

MuJoCoUnsupervised Reinforcement Learning

Behavior Contrastive Learning for Unsupervised Skill Discovery

2023-05-08 · Rushuai Yang, Chenjia Bai, Hongyi Guo, Siyuan Li 외

In reinforcement learning, unsupervised skill discovery aims to learn diverse skills without extrinsic rewards. Previous methods discover skills by maximizing the mutual information (MI) between states and skills. Howeve…

continuous-controlContinuous ControlContrastive Learning

Constructing Skill Trees for Reinforcement Learning Agents from Demonstration Trajectories

2010-12-01 · NeurIPS 2010 12 · George Konidaris, Scott Kuindersma, Roderic Grupen, Andrew G. Barto

We introduce CST, an algorithm for constructing skill trees from demonstration trajectories in continuous reinforcement learning domains. CST uses a changepoint detection method to segment each trajectory into a skill ch…

reinforcement-learningReinforcement LearningReinforcement Learning (RL)

Bayesian Nonparametrics for Offline Skill Discovery

2022-02-09 · Valentin Villecroze, Harry J. Braviner, Panteha Naderian, Chris J. Maddison 외

Skills or low-level policies in reinforcement learning are temporally extended actions that can speed up learning and enable complex behaviours. Recent work in offline reinforcement learning and imitation learning has pr…

Imitation Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)+1