A Robust Strategy for UAV Autonomous Landing on a Moving Platform under Partial Observability-Reference-Cited by-同舟云学术

A Robust Strategy for UAV Autonomous Landing on a Moving Platform under Partial Observability

Published:2024-05-30 Issue:6 Volume:8 Page:232
ISSN:2504-446X
Container-title:Drones
language:en
Short-container-title:Drones

Author:

Aikins Godwyll¹,Jagtap Sagar¹,Nguyen Kim-Doang¹^ORCID

Affiliation:

1. Department of Mechanical and Civil Engineering, Florida Institute of Technology, Melbourne, FL 32901, USA

Abstract

Landing a multi-rotor uncrewed aerial vehicle (UAV) on a moving target in the presence of partial observability, due to factors such as sensor failure or noise, represents an outstanding challenge that requires integrative techniques in robotics and machine learning. In this paper, we propose embedding a long short-term memory (LSTM) network into a variation of proximal policy optimization (PPO) architecture, termed robust policy optimization (RPO), to address this issue. The proposed algorithm is a deep reinforcement learning approach that utilizes recurrent neural networks (RNNs) as a memory component. Leveraging the end-to-end learning capability of deep reinforcement learning, the RPO-LSTM algorithm learns the optimal control policy without the need for feature engineering. Through a series of simulation-based studies, we demonstrate the superior effectiveness and practicality of our approach compared to the state-of-the-art proximal policy optimization (PPO) and the classical control method Lee-EKF, particularly in scenarios with partial observability. The empirical results reveal that RPO-LSTM significantly outperforms competing reinforcement learning algorithms, achieving up to 74% more successful landings than Lee-EKF and 50% more than PPO in flicker scenarios, maintaining robust performance in noisy environments and in the most challenging conditions that combine flicker and noise. These findings underscore the potential of RPO-LSTM in solving the problem of UAV landing on moving targets amid various degrees of sensor impairment and environmental interference.

Funder

National Science Foundation

Publisher

MDPI AG

Link

https://www.mdpi.com/2504-446X/8/6/232/pdf

Reference41 articles.

1. Optimal surveillance coverage for teams of micro aerial vehicles in GPS-denied environments using onboard vision;Doitsidis;Auton. Robot.,2012

2. Airborne Wind Energy Systems: A review of the technologies;Cherubini;Renew. Sustain. Energy Rev.,2015

3. Williams, A., and Yakimenko, O. (2018, January 20–23). Persistent mobile aerial surveillance platform using intelligent battery health management and drone swapping. Proceedings of the 2018 4th International Conference on Control, Automation and Robotics (ICCAR), Auckland, New Zealand.

4. Scott, J., and Scott, C. (2017, January 4–7). Drone delivery models for healthcare. Proceedings of the 50th Hawaii International Conference on System Sciences, Hilton Waikoloa Village, HI, USA.

5. Arora, S., Jain, S., Scherer, S., Nuske, S., Chamberlain, L., and Singh, S. (2013, January 6–10). Infrastructure-free shipdeck tracking for autonomous landing. Proceedings of the 2013 IEEE International Conference on Robotics and Automation, Karlsruhe, Germany.