Skip to main navigation Skip to search Skip to main content

基于多智能体强化学习的空间机械臂轨迹规划

Translated title of the contribution: Trajectory planning of space manipulator based on multi-agent reinforcement learning
  • School of Astronautics, Harbin Institute of Technology

Research output: Contribution to journalArticlepeer-review

Abstract

An online self-learning trajectory planning method based on the deep reinforcement learning is studied for a six Degree-of-Freedom (DOF) space floating manipulator to capture moving objects. The DH(Denavit-Hartenberg) model of the manipulator is presented, and the kinematic and dynamic models of multi-rigid bodies established considering the mechanical coupling characteristics of the combination. An improved deep determination policy gradient algorithm is further proposed, and a multi-agent self-learning system established with each joint as a decision-making agent. Additionally, a training model of the space manipulator is built based on "offline centralized learning and online distributed execution", constructing a reward function with the variables of the target relative distance and the total operation time. Simulation results show that the robot can capture the moving target rapidly with average time of 5.4 s. Compared with the traditional planning algorithm based on random sampling, the autonomous decision-making motion planning method proposed in this paper exhibits better solution speed and robustness.

Translated title of the contributionTrajectory planning of space manipulator based on multi-agent reinforcement learning
Original languageChinese (Traditional)
Article number524151
JournalHangkong Xuebao/Acta Aeronautica et Astronautica Sinica
Volume42
Issue number1
DOIs
StatePublished - 25 Jan 2021
Externally publishedYes

Fingerprint

Dive into the research topics of 'Trajectory planning of space manipulator based on multi-agent reinforcement learning'. Together they form a unique fingerprint.

Cite this