paper-with-me

홈 › Papers

Curious Meta-Controller: Adaptive Alternation between Model-Based and Model-Free Control in Deep Reinforcement Learning

2019-05-05 · Muhammad Burhan Hafez, Cornelius Weber, Matthias Kerzel, Stefan Wermter

Recent success in deep reinforcement learning for continuous control has been dominated by model-free approaches which, unlike model-based approaches, do not suffer from representational limitations in making assumptions about the world dynamics and model errors inevitable in complex domains. However, they require a lot of experiences compared to model-based approaches that are typically more sample-efficient. We propose to combine the benefits of the two approaches by presenting an integrated approach called Curious Meta-Controller. Our approach alternates adaptively between model-based and model-free control using a curiosity feedback based on the learning progress of a neural model of the dynamics in a learned latent space. We demonstrate that our approach can significantly improve the sample efficiency and achieve near-optimal performance on learning robotic reaching and grasping tasks from raw-pixel input in both dense and sparse reward settings.

📄 PDF Abstract BibTeX arXiv:1905.01718

Code (0)

등록된 구현이 없습니다.

Tasks

continuous-controlContinuous ControlDeep Reinforcement LearningmodelReinforcement Learning

Similar Papers 제목 키워드 기반

Adaptive-Control-Oriented Meta-Learning for Nonlinear Systems

2021-03-07 · Spencer M. Richards, Navid Azizan, Jean-Jacques Slotine, Marco Pavone

Real-time adaptation is imperative to the control of robots operating in complex, dynamic environments. Adaptive control laws can endow even nonlinear systems with good trajectory tracking performance, provided that any …

Meta-Learningregression

Control-oriented meta-learning

2022-04-14 · Spencer M. Richards, Navid Azizan, Jean-Jacques Slotine, Marco Pavone

Real-time adaptation is imperative to the control of robots operating in complex, dynamic environments. Adaptive control laws can endow even nonlinear systems with good trajectory tracking performance, provided that any …

Meta-Learningregression

MetaTune: Adjoint-based Meta-tuning via Robotic Differentiable Dynamics

2026-03-28 · Xiexin Peng, Bingheng Wang, Tao Zhang, Que Dong 외 arxiv

Disturbance observer-based control has shown promise in robustifying robotic systems against uncertainties. However, tuning such systems remains challenging due to the strong coupling between controller gains and observe…

Active World Model Learning with Progress Curiosity

2020-07-15 · Kuno Kim, Megumi Sano, Julian De Freitas, Nick Haber 외

World models are self-supervised predictive models of how the world evolves. Humans learn world models by curiously exploring their environment, in the process acquiring compact abstractions of high bandwidth sensory inp…

model

Metacontrol for Adaptive Imagination-Based Optimization

2017-05-07 · Jessica B. Hamrick, Andrew J. Ballard, Razvan Pascanu, Oriol Vinyals 외

Many machine learning systems are built to solve the hardest examples of a particular task, which often makes them large and expensive to run---especially with respect to the easier examples, which might require much les…

Decision MakingReinforcement Learning