Comparative analysis of model performance for predicting the customer of cafeteria using unstructured data

  • Kim, Seungsik; 
  • Gu, Nami; 
  • Moon, Jeongin; 
  • Kim, Keunwook; 
  • Hwang, Yeongeun; 
  • 외 1명
Citations

WEB OF SCIENCE

0
Citations

SCOPUS

0

초록

This study aimed to predict the number of meals served in a group cafeteria using machine learning methodology. Features of the menu were created through the Word2Vec methodology and clustering, and a stacking ensemble model was constructed using Random Forest, Gradient Boosting, and CatBoost as sub-models. Results showed that CatBoost had the best performance with the ensemble model showing an 8% improvement in performance. The study also found that the date variable had the greatest influence on the number of diners in a cafeteria, followed by menu characteristics and other variables. The implications of the study include the potential for machine learning methodology to improve predictive performance and reduce food waste, as well as the removal of subjective elements in menu classification. Limitations of the research include limited data cases and a weak model structure when new menus or foreign words are not included in the learning data. Future studies should aim to address these limitations.

키워드

cafeteria; ensemble model; ESG; food waste; machine learning; menu features; performance improvement; prediction; word embedding
제목
Comparative analysis of model performance for predicting the customer of cafeteria using unstructured data
저자
Kim, Seungsik; Gu, Nami; Moon, Jeongin; Kim, Keunwook; Hwang, Yeongeun; Lee, Kyeongjun
DOI
10.29220/CSAM.2023.30.5.485
발행일
2023-09
유형
Article
저널명
Communications for Statistical Applications and Methods
권
30
호
5
페이지
485 ~ 499