paper-with-me

홈 › Papers

Data-Efficient Learning of Feedback Policies from Image Pixels using Deep Dynamical Models

2015-10-08 · John-Alexander M. Assael, Niklas Wahlström, Thomas B. Schön, Marc Peter Deisenroth

Data-efficient reinforcement learning (RL) in continuous state-action spaces using very high-dimensional observations remains a key challenge in developing fully autonomous systems. We consider a particularly important instance of this challenge, the pixels-to-torques problem, where an RL agent learns a closed-loop control policy ("torques") from pixel information only. We introduce a data-efficient, model-based reinforcement learning algorithm that learns such a closed-loop policy directly from pixel information. The key ingredient is a deep dynamical model for learning a low-dimensional feature embedding of images jointly with a predictive model in this low-dimensional feature space. Joint learning is crucial for long-term predictions, which lie at the core of the adaptive nonlinear model predictive control strategy that we use for closed-loop control. Compared to state-of-the-art RL methods for continuous states and actions, our approach learns quickly, scales to high-dimensional state spaces, is lightweight and an important step toward fully autonomous end-to-end learning from pixels to torques.

📄 PDF Abstract BibTeX arXiv:1510.02173

Code (0)

등록된 구현이 없습니다.

Tasks

Model-based Reinforcement LearningModel Predictive Controlreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Similar Papers 제목 키워드 기반

FU-net: Multi-class Image Segmentation Using Feedback Weighted U-net

2020-04-28 · Mina Jafari, Ruizhe Li, Yue Xing, Dorothee Auer 외

In this paper, we present a generic deep convolutional neural network (DCNN) for multi-class image segmentation. It is based on a well-established supervised end-to-end DCNN model, known as U-net. U-net is firstly modifi…

Image SegmentationSegmentationSemantic Segmentation

Change Point Detection Approach for Online Control of Unknown Time Varying Dynamical Systems

2022-10-21 · Deepan Muthirayan, Ruijie Du, Yanning Shen, Pramod P. Khargonekar

We propose a novel change point detection approach for online learning control with full information feedback (state, disturbance, and cost feedback) for unknown time-varying dynamical systems. We show that our algorithm…

Change Point Detection

MAD: A Magnitude And Direction Policy Parametrization for Stability Constrained Reinforcement Learning

2025-04-03 · Luca Furieri, Sucheth Shenoy, Danilo Saccani, Andrea Martin 외

We introduce magnitude and direction (MAD) policies, a policy parameterization for reinforcement learning (RL) that preserves Lp closed-loop stability for nonlinear dynamical systems. Although complete in their ability t…

Reinforcement Learning (RL)

Current Implicit Policies May Not Eradicate COVID-19

2022-03-29 · Ali Jadbabaie, Arnab Sarker, Devavrat Shah

Successful predictive modeling of epidemics requires an understanding of the implicit feedback control strategies which are implemented by populations to modulate the spread of contagion. While this task of capturing end…

A Black-box Adversarial Attack Strategy with Adjustable Sparsity and Generalizability for Deep Image Classifiers

2020-04-24 · Arka Ghosh, Sankha Subhra Mullick, Shounak Datta, Swagatam Das 외

Constructing adversarial perturbations for deep neural networks is an important direction of research. Crafting image-dependent adversarial perturbations using white-box feedback has hitherto been the norm for such adver…

Adversarial Attack