A Penetration Method for UAV Based on Distributed Reinforcement Learning and Demonstrations-Reference-Cited by-同舟云学术

A Penetration Method for UAV Based on Distributed Reinforcement Learning and Demonstrations

Published:2023-03-27 Issue:4 Volume:7 Page:232
ISSN:2504-446X
Container-title:Drones
language:en
Short-container-title:Drones

Author:

Li Kexv¹,Wang Yue¹,Zhuang Xing¹^ORCID,Yin Hao¹^ORCID,Liu Xinyu¹,Li Hanyu¹

Affiliation:

1. School of Mechatronical Engineering, Beijing Institute of Technology, Beijing 100081, China

Abstract

The penetration of unmanned aerial vehicles (UAVs) is an essential and important link in modern warfare. Enhancing UAV’s ability of autonomous penetration through machine learning has become a research hotspot. However, the current generation of autonomous penetration strategies for UAVs faces the problem of excessive sample demand. To reduce the sample demand, this paper proposes a combination policy learning (CPL) algorithm that combines distributed reinforcement learning and demonstrations. Innovatively, the action of the CPL algorithm is jointly determined by the initial policy obtained from demonstrations and the target policy in the asynchronous advantage actor-critic network, thus retaining the guiding role of demonstrations in the initial training. In a complex and unknown dynamic environment, 1000 training experiments and 500 test experiments were conducted for the CPL algorithm and related baseline algorithms. The results show that the CPL algorithm has the smallest sample demand, the highest convergence efficiency, and the highest success rate of penetration among all the algorithms, and has strong robustness in dynamic environments.

Publisher

MDPI AG

Subject

Artificial Intelligence,Computer Science Applications,Aerospace Engineering,Information Systems,Control and Systems Engineering

Link

https://www.mdpi.com/2504-446X/7/4/232/pdf

Reference33 articles.

1. Anti-Interception Guidance for Hypersonic Glide Vehicle: A Deep Reinforcement Learning Approach;Jiang;Aerospace,2022

2. Proportional Navigation-Based Collision Avoidance for UAVs;Han;Int. J. Control Autom. Syst.,2009

3. Singh, L. (2004, January 16–19). Autonomous missile avoidance using nonlinear model predictive control. Proceedings of the AIAA Guidance Navigation, and Control Conference and Exhibit, Providence, RI, USA.

4. Gagnon, E., Rabbath, C., and Lauzon, M. (2005, January 15–18). Receding horizons with heading constraints for collision avoidance. Proceedings of the AIAA Guidance Navigation, and Control Conference and Exhibit, San Francisco, CA, USA.

5. Watanabe, Y., Calise, A., Johnson, E., and Evers, J. (2006, January 21–24). Minimum-effort guidance for vision-based collision avoidance. Proceedings of the AIAA Atmospheric Flight Mechanics Conference and Exhibit, Keystone, CO, USA.