Dynamically adjusting the k-values of the ATCS rule in a flexible flow shop scenario with reinforcement learning

23Citations
Citations of this article
52Readers
Mendeley users who have this article in their library.

This article is free to access.

Abstract

Given the fact that finding the optimal sequence in a flexible flow shop is usually an NP-hard problem, priority-based sequencing rules are applied in many real-world scenarios. In this contribution, an innovative reinforcement learning approach is used as a hyper-heuristic to dynamically adjust the k-values of the ATCS sequencing rule in a complex manufacturing scenario. For different product mixes as well as different utilisation levels, the reinforcement learning approach is trained and compared to the k-values found with an extensive simulation study. This contribution presents a human comprehensible hyper-heuristic, which is able to adjust the k-values to internal and external stimuli and can reduce the mean tardiness up to 5%.

Cite

CITATION STYLE

APA

Heger, J., & Voss, T. (2023). Dynamically adjusting the k-values of the ATCS rule in a flexible flow shop scenario with reinforcement learning. International Journal of Production Research, 61(1), 147–161. https://doi.org/10.1080/00207543.2021.1943762

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free