Deep Reinforcement Learning for Time Optimal Velocity Control using Prior Knowledge
Autonomous navigation has recently gained great interest in the field of reinforcement learning. However, little attention was given to the time optimal velocity control problem, i.e. controlling a vehicle such that it travels at the maximal speed without becoming dynamically unstable (roll-over or sliding). Time optimal velocity control can be solved numerically using existing methods that are based on optimal control and vehicle dynamics. In this paper, we use deep reinforcement learning to generate the time optimal velocity control. Furthermore, we use the numerical solution to further improve the performance of the reinforcement learner. It is shown that the reinforcement learner outperforms the numerically derived solution, and that the hybrid approach (combining learning with the numerical solution) speeds up the training process.
Code (0)
등록된 구현이 없습니다.
Tasks
Autonomous NavigationDeep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)Methods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Learning multiple gaits of quadruped robot using hierarchical reinforcement learning
There is a growing interest in learning a velocity command tracking controller of quadruped robot using reinforcement learning due to its robustness and scalability. However, a single policy, trained end-to-end, usually …
Hierarchical Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)Autonomous Satellite Docking via Adaptive Optimal Output Rregulation: A Reinforcement Learning Approach
This paper describes an online off-policy data-driven reinforcement learning based-algorithm to regulate and control the relative position of a deputy satellite in an autonomous satellite docking problem. The optimal con…
Positionreinforcement-learningReinforcement Learning (RL)Taming Lagrangian Chaos with Multi-Objective Reinforcement Learning
We consider the problem of two active particles in 2D complex flows with the multi-objective goals of minimizing both the dispersion rate and the energy consumption of the pair. We approach the problem by means of Multi …
Multi-Objective Reinforcement LearningQ-Learningreinforcement-learningReinforcement Learning+1Learning Efficient Navigation in Vortical Flow Fields
Efficient point-to-point navigation in the presence of a background flow field is important for robotic applications such as ocean surveying. In such applications, robots may only have knowledge of their immediate surrou…
reinforcement-learningReinforcement LearningReinforcement Learning (RL)Data-driven control of COVID-19 in buildings: a reinforcement-learning approach
In addition to its public health crisis, COVID-19 pandemic has led to the shutdown and closure of workplaces with an estimated total cost of more than $16 trillion. Given the long hours an average person spends in buildi…
reinforcement-learningReinforcement Learning (RL)