Reinforcement learning for pursuit and evasion of microswimmers at low Reynolds number
We consider a model of two competing microswimming agents engaged in a pursue-evasion task within a low-Reynolds-number environment. Agents can only perform simple maneuvers and sense hydrodynamic disturbances, which provide ambiguous (partial) information about the opponent's position and motion. We frame the problem as a zero-sum game: The pursuer has to capture the evader in the shortest time, while the evader aims at deferring capture as long as possible. We show that the agents, trained via adversarial reinforcement learning, are able to overcome partial observability by discovering increasingly complex sequences of moves and countermoves that outperform known heuristic strategies and exploit the hydrodynamic environment.
Code (0)
등록된 구현이 없습니다.
Tasks
Positionreinforcement-learningReinforcement Learning (RL)Similar Papers 제목 키워드 기반
Autonomous Decision Making for UAV Cooperative Pursuit-Evasion Game with Reinforcement Learning
The application of intelligent decision-making in unmanned aerial vehicle (UAV) is increasing, and with the development of UAV 1v1 pursuit-evasion game, multi-UAV cooperative game has emerged as a new challenge. This pap…
Decision MakingDeep Reinforcement Learningreinforcement-learningReinforcement LearningA Dynamics Perspective of Pursuit-Evasion Games of Intelligent Agents with the Ability to Learn
Pursuit-evasion games are ubiquitous in nature and in an artificial world. In nature, pursuer(s) and evader(s) are intelligent agents that can learn from experience, and dynamics (i.e., Newtonian or Lagrangian) is vital …
reinforcement-learningReinforcement LearningReinforcement Learning (RL)Online Planning for Multi-UAV Pursuit-Evasion in Unknown Environments Using Deep Reinforcement Learning
Multi-UAV pursuit-evasion, where pursuers aim to capture evaders, poses a key challenge for UAV swarm intelligence. Multi-agent reinforcement learning (MARL) has demonstrated potential in modeling cooperative behaviors, …
Deep Reinforcement LearningMulti-agent Reinforcement LearningThompson Sampling for Pursuit-Evasion Problems
Pursuit-evasion is a multi-agent sequential decision problem wherein a group of agents known as pursuers coordinate their traversal of a spatial domain to locate an agent trying to evade them. Pursuit evasion problems ar…
Thompson SamplingAdversary agent reinforcement learning for pursuit-evasion
A reinforcement learning environment with adversary agents is proposed in this work for pursuit-evasion game in the presence of fog of war, which is of both scientific significance and practical importance in aerospace a…
reinforcement-learningReinforcement LearningReinforcement Learning (RL)Starcraft