paper-with-me

Papers

Homotopy Based Reinforcement Learning with Maximum Entropy for Autonomous Air Combat

2021-12-01 · Yiwen Zhu, Zhou Fang, Yuan Zheng, Wenya Wei

The Intelligent decision of the unmanned combat aerial vehicle (UCAV) has long been a challenging problem. The conventional search method can hardly satisfy the real-time demand during high dynamics air combat scenarios. The reinforcement learning (RL) method can significantly shorten the decision time via using neural networks. However, the sparse reward problem limits its convergence speed and the artificial prior experience reward can easily deviate its optimal convergent direction of the original task, which raises great difficulties for the RL air combat application. In this paper, we propose a homotopy-based soft actor-critic method (HSAC) which focuses on addressing these problems via following the homotopy path between the original task with sparse reward and the auxiliary task with artificial prior experience reward. The convergence and the feasibility of this method are also proved in this paper. To confirm our method feasibly, we construct a detailed 3D air combat simulation environment for the RL-based methods training firstly, and we implement our method in both the attack horizontal flight UCAV task and the self-play confrontation task. Experimental results show that our method performs better than the methods only utilizing the sparse reward or the artificial prior experience reward. The agent trained by our method can reach more than 98.3% win rate in the attack horizontal flight UCAV task and average 67.4% win rate when confronted with the agents trained by the other two methods.

📄 PDF Abstract BibTeX arXiv:2112.01328

Code (0)

등록된 구현이 없습니다.

Tasks

reinforcement-learningReinforcement Learning (RL)

Methods 이 논문이 사용한 방법론

SPEED The monocular depth estimation (MDE) is the task of estimating depth from a single frame. This information is an essential knowledge in many computer vision tasks such as scene…

Similar Papers 제목 키워드 기반

Hierarchical Reinforcement Learning for Air-to-Air Combat

2021-05-03 · Adrian P. Pope, Jaime S. Ide, Daria Micovic, Henry Diaz 외

Artificial Intelligence (AI) is becoming a critical component in the defense industry, as recently demonstrated by DARPA`s AlphaDogfight Trials (ADT). ADT sought to vet the feasibility of AI algorithms capable of pilotin…

Hierarchical Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Active Perception in Adversarial Scenarios using Maximum Entropy Deep Reinforcement Learning

2019-02-14 · Macheng Shen, Jonathan P. How

We pose an active perception problem where an autonomous agent actively interacts with a second agent with potentially adversarial behaviors. Given the uncertainty in the intent of the other agent, the objective is to co…

Deep Reinforcement Learningreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Multi-Kernel Correntropy for Robust Learning

2019-05-24 · Badong Chen, Yuqing Xie, Xin Wang, Zejian yuan 외

As a novel similarity measure that is defined as the expectation of a kernel function between two random variables, correntropy has been successfully applied in robust machine learning and signal processing to combat lar…

Efficient Sampling-Based Maximum Entropy Inverse Reinforcement Learning with Application to Autonomous Driving

2020-06-22 · Zheng Wu, Liting Sun, Wei Zhan, Chenyu Yang 외

In the past decades, we have witnessed significant progress in the domain of autonomous driving. Advanced techniques based on optimization and reinforcement learning (RL) become increasingly powerful at solving the forwa…

Autonomous DrivingAutonomous Vehiclesreinforcement-learningReinforcement Learning+1

Maximum Entropy Diverse Exploration: Disentangling Maximum Entropy Reinforcement Learning

2019-11-03 · Andrew Cohen, Lei Yu, Xingye Qiao, Xiangrong Tong

Two hitherto disconnected threads of research, diverse exploration (DE) and maximum entropy RL have addressed a wide range of problems facing reinforcement learning algorithms via ostensibly distinct mechanisms. In this …

Diversityreinforcement-learningReinforcement LearningReinforcement Learning (RL)