Screening goals and selecting policies in hierarchical reinforcement learning

1Citations
Citations of this article
2Readers
Mendeley users who have this article in their library.
Get full text

Abstract

Hierarchical Reinforcement Learning (HRL) is primarily proposed for addressing problems with sparse reward signals and a long time horizon. Many existing HRL algorithms use neural networks to automatically produce goals, which have not taken into account that not all goals advance the task. Some methods have addressed the optimization of goal generation, while the goal is represented by the specific value of the state. In this paper, we propose a novel HRL algorithm for automatically discovering goals, which solves the problem for the optimization of goal generation in the latent state space by screening goals and selecting policies. We compare our approach with the state-of-the-art algorithms on Atari 2600 games and the results show that it can speed up learning and improve performance.

Cite

CITATION STYLE

APA

Zhou, J., Chen, J., Tong, Y., & Zhang, J. (2022). Screening goals and selecting policies in hierarchical reinforcement learning. Applied Intelligence, 52(15), 18049–18060. https://doi.org/10.1007/s10489-021-03093-9

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free