Deep Imitation Learning for Optimal Trajectory Planning and Initial Condition Optimization for an Unstable Dynamic System

Bo-Hsun Chen; Pei-Chun Lin

doi:10.1002/aisy.202300379

Advanced Intelligent Systems (Jan 2024)

Deep Imitation Learning for Optimal Trajectory Planning and Initial Condition Optimization for an Unstable Dynamic System

Bo-Hsun Chen,
Pei-Chun Lin

Affiliations

Bo-Hsun Chen: Department of Mechanical Engineering National Taiwan University (NTU) No.1 Roosevelt Rd. Sec.4 Taipei 106 Taiwan
Pei-Chun Lin: Department of Mechanical Engineering National Taiwan University (NTU) No.1 Roosevelt Rd. Sec.4 Taipei 106 Taiwan

DOI: https://doi.org/10.1002/aisy.202300379
Journal volume & issue: Vol. 6, no. 1
pp. n/a – n/a

Abstract

Read online

In this article, an innovative offline deep imitation learning algorithm for optimal trajectory planning is proposed. While many state‐of‐the‐art works achieved optimal trajectory planning, their systems were stable or quasistable, and their approaches rarely optimized the system's initial conditions (ICs). Here, a new unstable dynamic system task called “internal sliding object stabilization control” is proposed, modeled, and solved by deep imitation learning. Given the system's ICs, the neural networks (NNs) can imitate the iterative linear quadratic regulator (iLQR), generate optimal trajectories, and compute faster. A proportional–integral–derivative (PID) controller is used to track the unstable trajectories. Leveraging on the gradients of NNs, it can optimize the system's ICs, avoid obstacles stepwise, and ensure the worst bounds of NNs for safety. Subsequently, thorough simulations are conducted, including comparing the iLQR and PID controllers in the task, optimizing the system's different ICs by gradient descent, and finding the worst bound of the performance by gradient ascent. Results show that the proposed algorithm achieves considerably improved performance. Finally, experiments are conducted with a real manipulator to compare the proposed structure with the original iLQR. Results indicate that the proposed algorithm resembles the iLQR well. Program code and experiment results are in https://github.com/DanielYamChen/ISOSC.git.

Published in Advanced Intelligent Systems

ISSN: 2640-4567 (Online)
Publisher: Wiley
Country of publisher: Germany
LCC subjects: Technology: Electrical engineering. Electronics. Nuclear engineering: Electronics: Computer engineering. Computer hardware; Technology: Mechanical engineering and machinery: Control engineering systems. Automatic machinery (General)
Website: https://onlinelibrary.wiley.com/journal/26404567

About the journal

Abstract

Keywords