Abstract
Reinforcement learning (RL) is typically framed as a machine learning paradigm where agents learn to act autonomously in complex environments. This paper argues instead that RL is fundamentally human in the loop (HitL). The reward functions (and other components) of a Markov decision process are defined by humans. The decisions to tackle a certain problem, and deploy a learned solution, are taken by humans. Humans can also play a critical role in providing information to the agent throughout its life cycle to better succeed at the problem in question. We end by highlighting a set of critical HitL research questions, which, if ignored, could cause RL to fail to live up to its full potential.
Author supplied keywords
Cite
CITATION STYLE
Taylor, M. E. (2023). Reinforcement Learning Requires Human-in-the-Loop Framing and Approaches. In Frontiers in Artificial Intelligence and Applications (Vol. 368, pp. 351–360). IOS Press BV. https://doi.org/10.3233/FAIA230098
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.