Dashing for the Golden Snitch: Multi-Drone Time-Optimal Motion Planning with Multi-Agent Reinforcement Learning
Recent innovations in autonomous drones have facilitated time-optimal flight in single-drone configurations, and enhanced maneuverability in multi-drone systems by applying optimal control and learning-based methods. However, few studies have achieved time-optimal motion planning for multi-drone systems, particularly during highly agile maneuvers or in dynamic scenarios. This paper presents a decentralized policy network using multi-agent reinforcement learning for time-optimal multi-drone flight. To strike a balance between flight efficiency and collision avoidance, we introduce a soft collision-free mechanism inspired by optimization-based methods. By customizing PPO in a centralized training, decentralized execution (CTDE) fashion, we unlock higher efficiency and stability in training while ensuring lightweight implementation. Extensive simulations show that, despite slight performance trade-offs compared to single-drone systems, our multi-drone approach maintains near-time-optimal performance with a low collision rate. Real-world experiments validate our method, with two quadrotors using the same network as in simulation achieving a maximum speed of 13.65 m/s and a maximum body rate of 13.4 rad/s in a 5.5 m * 5.5 m * 2.0 m space across various tracks, relying entirely on onboard computation.
Code (1)
Tasks
Collision AvoidanceMotion PlanningMulti-agent Reinforcement LearningOptimal Motion PlanningMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
YOLO-Drone:Airborne real-time detection of dense small objects from high-altitude perspective
Unmanned Aerial Vehicles (UAVs), specifically drones equipped with remote sensing object detection technology, have rapidly gained a broad spectrum of applications and emerged as one of the primary research focuses in th…
Objectobject-detectionObject DetectionReal-Time Object DetectionHigh Definition image classification in Geoscience using Machine Learning
High Definition (HD) digital photos taken with drones are widely used in the study of Geoscience. However, blurry images are often taken in collected data, and it takes a lot of time and effort to distinguish clear image…
BIG-bench Machine LearningClassificationGeneral Classificationimage-classification+2Investigating Training Data Detection in AI Coders
Recent advances in code large language models (CodeLLMs) have made them indispensable tools in modern software engineering. However, these models occasionally produce outputs that contain proprietary or sensitive code sn…
Golden Ratio Information for Neural Spike Code
Spike bursting is a ubiquitous feature of all neuronal systems. Assuming the spiking states form an alphabet for a communication system, what is the optimal information processing rate? and what is the channel capacity? …
Communications and Control for Wireless Drone-Based Antenna Array
In this paper, the effective use of multiple quadrotor drones as an aerial antenna array that provides wireless service to ground users is investigated. In particular, under the goal of minimizing the airborne service ti…