Skip to main navigation Skip to search Skip to main content

再入飞行器深度确定性策略梯度制导方法研究

Translated title of the contribution: Research on deep deterministic policy gradient guidance method for reentry vehicle
  • Dongzi Guo
  • , Rong Huang
  • , Hechuan Xu
  • , Liwei Sun
  • , Naigang Cui*
  • *Corresponding author for this work
  • School of Astronautics, Harbin Institute of Technology
  • Beijing Institute of Control and Electronic Technology
  • NORINCO Group

Research output: Contribution to journalArticlepeer-review

Abstract

In order to solve the problem that the traditional reentry vehicle trajectory guidance methods are not adaptable to the strong disturbance conditions and difficult to meet the terminal constraints. Based on the framework of deep deterministic policy gradient (DDPG) reinforcement learning method, conducts network training on the off-line flight trajectory under the random strong disturbance conditions to find the optimal actor network under different environmental conditions. It can be used for guidance trajectory planning under the condition of on-line interference to meet the terminal altitude, range and speed constraints of reentry flight by periodically forecasting the angle of attack and pitch profile of reentry flight. The simulation results show that the maximum terminal residual range deviation is less than 500 m and the maximum terminal speed deviation is less than 35 m/s while meeting the terminal height constraint. Compared with the traditional tracking guidance method, the guidance control method proposed in this paper has higher accuracy and less calculation, which has a good engineering application prospect.

Translated title of the contributionResearch on deep deterministic policy gradient guidance method for reentry vehicle
Original languageChinese (Traditional)
Pages (from-to)1942-1949
Number of pages8
JournalXi Tong Gong Cheng Yu Dian Zi Ji Shu/Systems Engineering and Electronics
Volume44
Issue number6
DOIs
StatePublished - Jun 2022
Externally publishedYes

Fingerprint

Dive into the research topics of 'Research on deep deterministic policy gradient guidance method for reentry vehicle'. Together they form a unique fingerprint.

Cite this