Velocity Control of a Multi-Motion Mode Spherical Probe Robot Based on Reinforcement Learning

Wenke Ma; Bingyang Li; Yuxue Cao; Pengfei Wang; Mengyue Liu; Chenyang Chang; Shigang Peng

doi:10.3390/app13148218

Applied Sciences (Jul 2023)

Velocity Control of a Multi-Motion Mode Spherical Probe Robot Based on Reinforcement Learning

Wenke Ma,
Bingyang Li,
Yuxue Cao,
Pengfei Wang,
Mengyue Liu,
Chenyang Chang,
Shigang Peng

Affiliations

Wenke Ma: Qian Xuesen Laboratory of Space Technology, China Academy of Space Technology, Beijing 100094, China
Bingyang Li: China Academy of Aerospace Science and Innovation, Beijing 102600, China
Yuxue Cao: Beijing Institute of Control Engineering, Beijing 100190, China
Pengfei Wang: China Academy of Aerospace Science and Innovation, Beijing 102600, China
Mengyue Liu: China Academy of Aerospace Science and Innovation, Beijing 102600, China
Chenyang Chang: College of Engineering, Peking University, Beijing 100871, China
Shigang Peng: Qian Xuesen Laboratory of Space Technology, China Academy of Space Technology, Beijing 100094, China

DOI: https://doi.org/10.3390/app13148218
Journal volume & issue: Vol. 13, no. 14
p. 8218

Abstract

Read online

As deep space exploration tasks become increasingly complex, the mobility and adaptability of traditional wheeled or tracked probe robots with high functional density are constrained in harsh, dangerous, or unknown environments. A practical solution to these challenges is designing a probe robot for preliminary exploration in unknown areas, which is characterized by robust adaptability, simple structure, light weight, and minimal volume. Compared to the traditional deep space probe robot, the spherical robot with a geometric, symmetrical structure shows better adaptability to the complex ground environment. Considering the uncertain detection environment, the spherical robot should brake rapidly after jumping to avoid reentering obstacles. Moreover, since it is equipped with optical modules for deep space exploration missions, the spherical robot must maintain motion stability during the rolling process to ensure the quality of photos and videos captured. However, due to the nonlinear coupling and parameter uncertainty of the spherical robot, it is tedious to adjust controller parameters. Moreover, the adaptability of controllers with fixed parameters is limited. This paper proposes an adaptive proportion–integration–differentiation (PID) control method based on reinforcement learning for the multi-motion mode spherical probe robot (MMSPR) with rolling and jumping. This method uses the soft actor–critic (SAC) algorithm to adjust the parameters of the PID controller and introduces a switching control strategy to reduce static error. As the simulation results show, this method can facilitate the MMSPR’s convergence within 0.02 s regarding motion stability. In addition, in terms of braking, it enables an MMSPR with random initial speed brake within a convergence time of 0.045 s and a displacement of 0.0013 m. Compared with the PID method with fixed parameters, the braking displacement of the MMSPR is reduced by about 38%, and the convergence time is reduced by about 20%, showing better universality and adaptability.

Published in Applied Sciences

ISSN: 2076-3417 (Online)
Publisher: MDPI AG
Country of publisher: Switzerland
LCC subjects: Technology: Engineering (General). Civil engineering (General); Science: Biology (General); Science: Physics; Science: Chemistry
Website: http://www.mdpi.com/journal/applsci

About the journal

Abstract

Keywords