Research on autonomous collision avoidance of merchant ship based on inverse reinforcement learning

22Citations
Citations of this article
31Readers
Mendeley users who have this article in their library.

This article is free to access.

Abstract

To learn the optimal collision avoidance policy of merchant ships controlled by human experts, a finite-state Markov decision process model for ship collision avoidance is proposed based on the analysis of collision avoidance mechanism, and an inverse reinforcement learning (IRL) method based on cross entropy and projection is proposed to obtain the optimal policy from expert’s demonstrations. Collision avoidance simulations in different ship encounters are conducted and the results show that the policy obtained by the proposed IRL has a good inversion effect on two kinds of human experts, which indicate that the proposed method can effectively learn the policy of human experts for ship collision avoidance.

Cite

CITATION STYLE

APA

Zheng, M., Xie, S., Chu, X., Zhu, T., & Tian, G. (2020). Research on autonomous collision avoidance of merchant ship based on inverse reinforcement learning. International Journal of Advanced Robotic Systems, 17(6). https://doi.org/10.1177/1729881420969081

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free