Deep Imitation Learning for Optimal Trajectory Planning and Initial Condition Optimization for an Unstable Dynamic System-Reference-Cited by-同舟云学术

Deep Imitation Learning for Optimal Trajectory Planning and Initial Condition Optimization for an Unstable Dynamic System

Published:2023-11-27 Issue:1 Volume:6 Page:
ISSN:2640-4567
Container-title:Advanced Intelligent Systems
language:en
Short-container-title:Advanced Intelligent Systems

Author:

Chen Bo-Hsun¹,Lin Pei-Chun¹^ORCID

Affiliation:

1. Department of Mechanical Engineering National Taiwan University (NTU) No.1 Roosevelt Rd. Sec.4 Taipei 106 Taiwan

Abstract

In this article, an innovative offline deep imitation learning algorithm for optimal trajectory planning is proposed. While many state‐of‐the‐art works achieved optimal trajectory planning, their systems were stable or quasistable, and their approaches rarely optimized the system's initial conditions (ICs). Here, a new unstable dynamic system task called “internal sliding object stabilization control” is proposed, modeled, and solved by deep imitation learning. Given the system's ICs, the neural networks (NNs) can imitate the iterative linear quadratic regulator (iLQR), generate optimal trajectories, and compute faster. A proportional–integral–derivative (PID) controller is used to track the unstable trajectories. Leveraging on the gradients of NNs, it can optimize the system's ICs, avoid obstacles stepwise, and ensure the worst bounds of NNs for safety. Subsequently, thorough simulations are conducted, including comparing the iLQR and PID controllers in the task, optimizing the system's different ICs by gradient descent, and finding the worst bound of the performance by gradient ascent. Results show that the proposed algorithm achieves considerably improved performance. Finally, experiments are conducted with a real manipulator to compare the proposed structure with the original iLQR. Results indicate that the proposed algorithm resembles the iLQR well. Program code and experiment results are in https://github.com/DanielYamChen/ISOSC.git.

Funder

National Science and Technology Council

Publisher

Wiley

Subject

General Medicine

Link

https://onlinelibrary.wiley.com/doi/pdf/10.1002/aisy.202300379

Reference40 articles.

1. Reinforcement Learning-Based Collision Avoidance and Optimal Trajectory Planning in UAV Communication Networks

2. Nonlinear Model Predictive Horizon for Optimal Trajectory Generation

3. Time-Optimal Trajectory Planning With Interaction With the Environment

4. CHOMP: Covariant Hamiltonian optimization for motion planning

5. Motion planning with sequential convex optimization and convex collision checking