An AUV Target-Tracking Method Combining Imitation Learning and Deep Reinforcement Learning

23Citations
Citations of this article
18Readers
Mendeley users who have this article in their library.

Abstract

This study aims to solve the problem of sparse reward and local convergence when using a reinforcement learning algorithm as the controller of an AUV. Based on the generative adversarial imitation (GAIL) algorithm combined with a multi-agent, a multi-agent GAIL (MAG) algorithm is proposed. The GAIL enables the AUV to directly learn from expert demonstrations, overcoming the difficulty of slow initial training of the network. Parallel training of multi-agents reduces the high correlation between samples to avoid local convergence. In addition, a reward function is designed to help training. Finally, the results show that in the unity simulation platform test, the proposed algorithm has a strong optimal decision-making ability in the tracking process.

Cite

CITATION STYLE

APA

Mao, Y., Gao, F., Zhang, Q., & Yang, Z. (2022). An AUV Target-Tracking Method Combining Imitation Learning and Deep Reinforcement Learning. Journal of Marine Science and Engineering, 10(3). https://doi.org/10.3390/jmse10030383

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free