A guidance method for coplanar orbital interception based on reinforcement learning

Zeng Xin; Zhu Yanwei; Yang Leping; Zhang Chengming

Journal ArticleOPEN ACCESS

A guidance method for coplanar orbital interception based on reinforcement learning

Journal of Systems Engineering and Electronics (2021) 32(4) 927-938

DOI: 10.23919/JSEE.2021.000079

15Citations

5Readers

Abstract

This paper investigates the guidance method based on reinforcement learning (RL) for the coplanar orbital interception in a continuous low-thrust scenario. The problem is formulated into a Markov decision process (MDP) model, then a well-designed RL algorithm, experience based deep deterministic policy gradient (EBDDPG), is proposed to solve it. By taking the advantage of prior information generated through the optimal control model, the proposed algorithm not only resolves the convergence problem of the common RL algorithm, but also successfully trains an efficient deep neural network (DNN) controller for the chaser spacecraft to generate the control sequence. Numerical simulation results show that the proposed algorithm is feasible and the trained DNN controller significantly improves the efficiency over traditional optimization methods by roughly two orders of magnitude.

Author supplied keywords

Cite

CITATION STYLE

APA

Xin, Z., Yanwei, Z., Leping, Y., & Chengming, Z. (2021). A guidance method for coplanar orbital interception based on reinforcement learning. Journal of Systems Engineering and Electronics, 32(4), 927–938. https://doi.org/10.23919/JSEE.2021.000079

A guidance method for coplanar orbital interception based on reinforcement learning

Abstract

Author supplied keywords

Cite

Register to see more suggestions