Multi-UAV Path Planning and Following Based on Multi-Agent Reinforcement Learning-Reference-Cited by-同舟云学术

Multi-UAV Path Planning and Following Based on Multi-Agent Reinforcement Learning

Published:2024-01-11 Issue:1 Volume:8 Page:18
ISSN:2504-446X
Container-title:Drones
language:en
Short-container-title:Drones

Author:

Zhao Xiaoru¹^ORCID,Yang Rennong¹,Zhong Liangsheng²^ORCID,Hou Zhiwei²

Affiliation:

1. Air Traffic Control and Navigation School, Air Force Engineering University, Xi’an 710051, China

2. School of Systems Science and Engineering, Sun Yat-sen University, Guangzhou 510275, China

Abstract

Dedicated to meeting the growing demand for multi-agent collaboration in complex scenarios, this paper introduces a parameter-sharing off-policy multi-agent path planning and the following approach. Current multi-agent path planning predominantly relies on grid-based maps, whereas our proposed approach utilizes laser scan data as input, providing a closer simulation of real-world applications. In this approach, the unmanned aerial vehicle (UAV) uses the soft actor–critic (SAC) algorithm as a planner and trains its policy to converge. This policy enables end-to-end processing of laser scan data, guiding the UAV to avoid obstacles and reach the goal. At the same time, the planner incorporates paths generated by a sampling-based method as following points. The following points are continuously updated as the UAV progresses. Multi-UAV path planning tasks are facilitated, and policy convergence is accelerated through sharing experiences among agents. To address the challenge of UAVs that are initially stationary and overly cautious near the goal, a reward function is designed to encourage UAV movement. Additionally, a multi-UAV simulation environment is established to simulate real-world UAV scenarios to support training and validation of the proposed approach. The simulation results highlight the effectiveness of the presented approach in both the training process and task performance. The presented algorithm achieves an 80% success rate to guarantee that three UAVs reach the goal points.

Publisher

MDPI AG

Link

https://www.mdpi.com/2504-446X/8/1/18/pdf

Reference50 articles.

1. Madridano, Á., Al-Kaff, A., Gómez, D.M., and de la Escalera, A. (2019, January 4–6). Multi-Path Planning Method for UAVs Swarm Purposes. Proceedings of the 2019 IEEE International Conference on Vehicular Electronics and Safety (ICVES), Cairo, Egypt.

2. Lin, S., Liu, A., Wang, J., and Kong, X. (2022). A Review of Path-Planning Approaches for Multiple Mobile Robots. Machines, 10.

3. Distributed multi-robot collision avoidance via deep reinforcement learning for navigation in complex scenarios;Fan;Int. J. Robot. Res.,2020

4. UAV Path Planning Using Optimization Approaches: A Survey;Soukane;Arch. Comput. Methods Eng.,2022

5. Mechali, O., Xu, L., Wei, M., Benkhaddra, I., Guo, F., and Senouci, A. (August, January 29). A Rectified RRT* with Efficient Obstacles Avoidance Method for UAV in 3D Environment. Proceedings of the 2019 IEEE 9th Annual International Conference on CYBER Technology in Automation, Control, and Intelligent Systems (CYBER), Suzhou, China.

Cited by 5 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Energy-Efficient Online Path Planning for Internet of Drones Using Reinforcement Learning;Journal of Sensor and Actuator Networks;2024-08-29

2. Cloud-based Speech Recognition for UAV Control Architecture in Industry 4.0;2024 International Conference on Artificial Intelligence, Big Data, Computing and Data Communication Systems (icABCD);2024-08-01

3. Simulation Training System for Parafoil Motion Controller Based on Actor–Critic RL Approach;Actuators;2024-07-25

4. Collaborative Encirclement of Multiple UAVs Based on Deep Reinforcement Learning;2024 36th Chinese Control and Decision Conference (CCDC);2024-05-25

5. IR-QLA: Machine Learning-Based Q-Learning Algorithm Optimization for UAVs Faster Trajectory Planning by Instructed- Reinforcement Learning;IEEE Access;2024