A Human Mixed Strategy Approach to Deep Reinforcement Learning
In 2015, Google's DeepMind announced an advancement in creating an autonomous agent based on deep reinforcement learning (DRL) that could beat a professional player in a series of 49 Atari games. However, the current manifestation of DRL is still immature, and has significant drawbacks. One of DRL's imperfections is its lack of "exploration" during the training process, especially when working with high-dimensional problems. In this paper, we propose a mixed strategy approach that mimics behaviors of human when interacting with environment, and create a "thinking" agent that allows for more efficient exploration in the DRL training process. The simulation results based on the Breakout game show that our scheme achieves a higher probability of obtaining a maximum score than does the baseline DRL algorithm, i.e., the asynchronous advantage actor-critic method. The proposed scheme therefore can be applied effectively to solving a complicated task in a real-world application.
Code (0)
등록된 구현이 없습니다.
Tasks
Atari GamesDeep Reinforcement LearningEfficient Explorationreinforcement-learningReinforcement LearningReinforcement Learning (RL)Similar Papers 제목 키워드 기반
Partially Connected Automated Vehicle Cooperative Control Strategy with a Deep Reinforcement Learning Approach
This paper proposes a cooperative strategy of connected and automated vehicles (CAVs) longitudinal control for partially connected and automated traffic environment based on deep reinforcement learning (DRL) algorithm, w…
Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)Optimal Weight Adaptation of Model Predictive Control for Connected and Automated Vehicles in Mixed Traffic with Bayesian Optimization
In this paper, we develop an optimal weight adaptation strategy of model predictive control (MPC) for connected and automated vehicles (CAVs) in mixed traffic. We model the interaction between a CAV and a human-driven ve…
Bayesian OptimizationModel Predictive ControlPlatooning Connected, Autonomous, and Human-Driven Vehicles: A Deep Reinforcement Learning-based Approach
Conventionally, existing vehicle platooning approaches are designed for connected vehicles, typically including connected autonomous vehicles and connected human-driven vehicles. Non-connected vehicles, such as non-conne…
Reinforcement LearningAutonomous VehiclesEnhancing System-Level Safety in Mixed-Autonomy Platoon via Safe Reinforcement Learning
Connected and automated vehicles (CAVs) have recently gained prominence in traffic research due to advances in communication technology and autonomous driving. Various longitudinal control strategies for CAVs have been d…
Autonomous DrivingCollision AvoidanceDeep Reinforcement LearningSafe Reinforcement LearningEnforcing Cooperative Safety for Reinforcement Learning-based Mixed-Autonomy Platoon Control
It is recognized that the control of mixed-autonomy platoons comprising connected and automated vehicles (CAVs) and human-driven vehicles (HDVs) can enhance traffic flow. Among existing methods, Multi-Agent Reinforcement…
Conformal PredictionMulti-agent Reinforcement Learning