A multi-agent cooperative reinforcement learning model using a hierarchy of consultants, tutors and workers

  • Abed-alguni B
  • Chalup S
  • Henskens F
  • et al.
N/ACitations
Citations of this article
32Readers
Mendeley users who have this article in their library.

This article is free to access.

Abstract

The hierarchical organisation of distributed sys-tems can provide an efficient decomposition for machine learning. This paper proposes an algorithm for coopera-tive policy construction for independent learners, named Q-learning with aggregation (QA-learning). The algorithm is based on a distributed hierarchical learning model and utilises three specialisations of agents: workers, tutors and consul-tants. The consultant agent incorporates the entire system in its problem space, which it decomposes into sub-problems that are assigned to the tutor and worker agents. The QA-learning algorithm aggregates the Q-tables of worker agents into a central repository managed by their tutor agent. Each tutor's Q-table is then incorporated into the consultant's Q-table, resulting in a Q-table for the entire problem. The algorithm was tested using a distributed hunter prey problem, and experimental results show that QA-learning converges to a solution faster than single agent Q-learning and some famous cooperative Q-learning algorithms.

Cite

CITATION STYLE

APA

Abed-alguni, B. H., Chalup, S. K., Henskens, F. A., & Paul, D. J. (2015). A multi-agent cooperative reinforcement learning model using a hierarchy of consultants, tutors and workers. Vietnam Journal of Computer Science, 2(4), 213–226. https://doi.org/10.1007/s40595-015-0045-x

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free