Enhancing UAV Aerial Docking: A Hybrid Approach Combining Offline and Online Reinforcement Learning-Reference-Cited by-同舟云学术

Enhancing UAV Aerial Docking: A Hybrid Approach Combining Offline and Online Reinforcement Learning

Published:2024-04-24 Issue:5 Volume:8 Page:168
ISSN:2504-446X
Container-title:Drones
language:en
Short-container-title:Drones

Author:

Feng Yuting¹^ORCID,Yang Tao¹,Yu Yushu¹

Affiliation:

1. The School of Mechatronical Engineering, Beijing Institute of Technology, Beijing 100081, China

Abstract

In our study, we explore the task of performing docking maneuvers between two unmanned aerial vehicles (UAVs) using a combination of offline and online reinforcement learning (RL) methods. This task requires a UAV to accomplish external docking while maintaining stable flight control, representing two distinct types of objectives at the task execution level. Direct online RL training could lead to catastrophic forgetting, resulting in training failure. To overcome these challenges, we design a rule-based expert controller and accumulate an extensive dataset. Based on this, we concurrently design a series of rewards and train a guiding policy through offline RL. Then, we conduct comparative verification on different RL methods, ultimately selecting online RL to fine-tune the model trained offline. This strategy effectively combines the efficiency of offline RL with the exploratory capabilities of online RL. Our approach improves the success rate of the UAV’s aerial docking task, increasing it from 40% under the expert policy to 95%.

Funder

National Natural Science Foundation of China

National Key R. D. Program of China

Publisher

MDPI AG

Link

https://www.mdpi.com/2504-446X/8/5/168/pdf

Reference55 articles.

1. Shot type constraints in UAV cinematography for autonomous target tracking;Karakostas;Inf. Sci.,2020

2. Real-Time Multi-Modal Active Vision for Object Detection on UAVs Equipped with Limited Field of View LiDAR and Camera;Shi;IEEE Robot. Autom. Lett.,2023

3. Communication and networking technologies for UAVs: A survey;Sharma;J. Netw. Comput. Appl.,2020

4. Design and Trajectory Linearization Geometric Control of Multiple Aerial Vehicles Assembly;Yu;J. Mech. Eng.,2022

5. A novel robotic platform for aerial manipulation using quadrotors as rotating thrust generators;Nguyen;IEEE Trans. Robot.,2018