paper-with-me

홈 › Papers

Learnings Options End-to-End for Continuous Action Tasks

2017-11-30 · Martin Klissarov, Pierre-Luc Bacon, Jean Harb, Doina Precup

We present new results on learning temporally extended actions for continuoustasks, using the options framework (Suttonet al.[1999b], Precup [2000]). In orderto achieve this goal we work with the option-critic architecture (Baconet al.[2017])using a deliberation cost and train it with proximal policy optimization (Schulmanet al.[2017]) instead of vanilla policy gradient. Results on Mujoco domains arepromising, but lead to interesting questions aboutwhena given option should beused, an issue directly connected to the use of initiation sets.

📄 PDF Abstract BibTeX arXiv:1712.00004

Code (3)

mklissa/PPOC 공식 구현 tf
kkhetarpal/ioc tf
shuishida/soaprl pytorch

Tasks

MuJoCo

Similar Papers 제목 키워드 기반

Dynamic Decision Frequency with Continuous Options

2022-12-06 · Amirmohammad Karimi, Jun Jin, Jun Luo, A. Rupam Mahmood 외

In classic reinforcement learning algorithms, agents make decisions at discrete and fixed time intervals. The duration between decisions becomes a crucial hyperparameter, as setting it too short may increase the problem'…

continuous-controlContinuous Control

Dynamically Addressing Unseen Rumor via Continual Learning

2021-04-18 · Nayeon Lee, Andrea Madotto, Yejin Bang, Pascale Fung

Rumors are often associated with newly emerging events, thus, an ability to deal with unseen rumors is crucial for a rumor veracity classification model. Previous works address this issue by improving the model's general…

Continual LearningVeracity Classification

Soft Options Critic

2019-05-23 · Elita Lobo, Scott Jordan

The option-critic architecture (Bacon, Harb, and Precup 2017) and several variants have successfully demonstrated the use of the options framework proposed by Sutton et al (Sutton, Precup, and Singh1999) to scale learnin…

Learning to Drive Safely with Hybrid Options

2025-10-28 · Bram De Cooman, Johan Suykens arxiv

Out of the many deep reinforcement learning approaches for autonomous driving, only few make use of the options (or skills) framework. That is surprising, as this framework is naturally suited for hierarchical control ap…

Reinforcement LearningAutonomous Driving

MO2: Model-Based Offline Options

2022-09-05 · Sasha Salter, Markus Wulfmeier, Dhruva Tirumala, Nicolas Heess 외

The ability to discover useful behaviours from past experience and transfer them to new tasks is considered a core component of natural embodied intelligence. Inspired by neuroscience, discovering behaviours that switch …

continuous-controlContinuous Controlmodel