Model-based micro-data reinforcement learning: what are the crucial model properties and which model to choose?
We contribute to micro-data model-based reinforcement learning (MBRL) by rigorously comparing popular generative models using a fixed (random shooting) control agent. We find that on an environment that requires multimodal posterior predictives, mixture density nets outperform all other models by a large margin. When multimodality is not required, our surprising finding is that we do not need probabilistic posterior predictives: deterministic models are on par, in fact they consistently (although non-significantly) outperform their probabilistic counterparts. We also found that heteroscedasticity at training time, perhaps acting as a regularizer, improves predictions at longer horizons. At the methodological side, we design metrics and an experimental protocol which can be used to evaluate the various models, predicting their asymptotic performance when using them on the control problem. Using this framework, we improve the state-of-the-art sample complexity of MBRL on Acrobot by two to four folds, using an aggressive training schedule which is outside of the hyperparameter interval usually considered
Code (1)
Tasks
AcrobotmodelModel-based Reinforcement LearningSimilar Papers 제목 키워드 기반
Intelligent Monitoring Framework for Cloud Services: A Data-Driven Approach
Cloud service owners need to continuously monitor their services to ensure high availability and reliability. Gaps in monitoring can lead to delay in incident detection and significant negative customer impact. Current p…
Optimal Control of Material Micro-Structures
In this paper, we consider the optimal control of material micro-structures. Such material micro-structures are modeled by the so-called phase field model. We study the underlying physical structure of the model and prop…
reinforcement-learningReinforcement LearningReinforcement Learning (RL)Lifelong Control of Off-grid Microgrid with Model Based Reinforcement Learning
The lifelong control problem of an off-grid microgrid is composed of two tasks, namely estimation of the condition of the microgrid devices and operational planning accounting for the uncertainties by forecasting the fut…
Model-based Reinforcement Learningreinforcement-learningReinforcement Learning (RL)ASQ-IT: Interactive Explanations for Reinforcement-Learning Agents
As reinforcement learning methods increasingly amass accomplishments, the need for comprehending their solutions becomes more crucial. Most explainable reinforcement learning (XRL) methods generate a static explanation d…
reinforcement-learningReinforcement LearningReinforcement Learning (RL)Multi-Objective Reinforcement Learning based Multi-Microgrid System Optimisation Problem
Microgrids with energy storage systems and distributed renewable energy sources play a crucial role in reducing the consumption from traditional power sources and the emission of $CO_2$. Connecting multi microgrid to a d…
energy managementManagementMulti-Objective Reinforcement Learningreinforcement-learning+2