Abstract
Enabling reinforcement learning (RL) agents to leverage a knowledge base while learning from experience promises to advance RL in knowledge intensive domains. However, it has proven difficult to leverage knowledge that is not manually tailored to the environment. We propose to use the subclass relationships present in open-source knowledge graphs to abstract away from specific objects. We develop a residual policy gradient method that is able to integrate knowledge across different abstraction levels in the class hierarchy. Our method results in improved sample efficiency and generalisation to unseen objects in commonsense games, but we also investigate failure modes, such as excessive noise in the extracted class knowledge or environments with little class structure.
Cite
CITATION STYLE
Höpner, N., Tiddi, I., & van Hoof, H. (2022). Leveraging Class Abstraction for Commonsense Reinforcement Learning via Residual Policy Gradient Methods. In IJCAI International Joint Conference on Artificial Intelligence (pp. 3050–3056). International Joint Conferences on Artificial Intelligence. https://doi.org/10.24963/ijcai.2022/423
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.