Abstract
This study aimed to predict the number of meals served in a group cafeteria using machine learning methodology. Features of the menu were created through the Word2Vec methodology and clustering, and a stacking ensemble model was constructed using Random Forest, Gradient Boosting, and CatBoost as sub-models. Results showed that CatBoost had the best performance with the ensemble model showing an 8% improvement in performance. The study also found that the date variable had the greatest influence on the number of diners in a cafeteria, followed by menu characteristics and other variables. The implications of the study include the potential for machine learning methodology to improve predictive performance and reduce food waste, as well as the removal of subjective elements in menu classification. Limitations of the research include limited data cases and a weak model structure when new menus or foreign words are not included in the learning data. Future studies should aim to address these limitations.
Author supplied keywords
Cite
CITATION STYLE
Kim, S., Gu, N., Moon, J., Kim, K., Hwang, Y., & Lee, K. (2023). Comparative analysis of model performance for predicting the customer of cafeteria using unstructured data. Communications for Statistical Applications and Methods, 30(5), 485–499. https://doi.org/10.29220/CSAM.2023.30.5.485
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.