Deep Reinforcement Learning for Weapons to Targets Assignment in a Hypersonic strike
We use deep reinforcement learning (RL) to optimize a weapons to target assignment (WTA) policy for multi-vehicle hypersonic strike against multiple targets. The objective is to maximize the total value of destroyed targets in each episode. Each randomly generated episode varies the number and initial conditions of the hypersonic strike weapons (HSW) and targets, the value distribution of the targets, and the probability of a HSW being intercepted. We compare the performance of this WTA policy to that of a benchmark WTA policy derived using non-linear integer programming (NLIP), and find that the RL WTA policy gives near optimal performance with a 1000X speedup in computation time, allowing real time operation that facilitates autonomous decision making in the mission end game.
Code (0)
등록된 구현이 없습니다.
Tasks
Decision MakingDeep Reinforcement Learningreinforcement-learningReinforcement Learning (RL)Similar Papers 제목 키워드 기반
Terminal Adaptive Guidance for Autonomous Hypersonic Strike Weapons via Reinforcement Learning
An adaptive guidance system suitable for the terminal phase trajectory of a hypersonic strike weapon is optimized using reinforcement meta learning. The guidance system maps observations directly to commanded bank angle,…
Meta-Learningreinforcement-learningReinforcement Learning (RL)ISAR Imaging Analysis of a Hypersonic Vehicle Covered With Plasma Sheath
In this article, a hypersonic target electromagnetic (EM) scattering echo model combined with the inhomogeneous zonal medium model (IZMM) and the classical scattering center model (SCM) is proposed with a distributed …
Motion CompensationHypersonic Flow Control: Generalized Deep Reinforcement Learning for Hypersonic Intake Unstart Control under Uncertainty
The hypersonic unstart phenomenon poses a major challenge to reliable air-breathing propulsion at Mach 5 and above, where strong shock-boundary-layer interactions and rapid pressure fluctuations can destabilize inlet ope…
Zero-shot GeneralizationReinforcement LearningSequence Compression Speeds Up Credit Assignment in Reinforcement Learning
Temporal credit assignment in reinforcement learning is challenging due to delayed and stochastic outcomes. Monte Carlo targets can bridge long delays between action and consequence but lead to high-variance targets due …
ChunkingNavigatereinforcement-learningReinforcement LearningFeedback Strategies for Hypersonic Pursuit of a Ground Evader
In this paper, we present a game-theoretic feedback terminal guidance law for an autonomous, unpowered hypersonic pursuit vehicle that seeks to intercept an evading ground target whose motion is constrained in a one-dime…