A Comparative Study of Machine Learning with Statistical Feature Selection for Risk Detection of Diabetic

  • Al Ghozali I
  • Fathin M
  • Handoko A
N/ACitations
Citations of this article
8Readers
Mendeley users who have this article in their library.

Abstract

Elevated glucose levels in the circulation are indicative of diabetes, a chronic medical condition. Prolonged unregulated blood glucose levels pose a significant risk of severe consequences, including renal failure, myocardial infarction, and lower limb amputation. The objective of this study is to conduct a comparative analysis of SVM, Naive Bayes, XGBoost, Random Forest, and ANN models in order to forecast the occurrence of diabetes. The research methodology comprises seven primary stages: (1) literature review, (2) data collection, (3) exploratory data analysis (EDA), (4) data preprocessing, (5) feature selection, (6) model development, and (7) model evaluation and comparison. The XGBoost model is the most suitable option, as indicated by the model evaluation results. The XGBoost model achieved a precision of 0.88, a recall of 0.87, and an accuracy of 0.8690. The XGBoost model has a RMSE of 0.3620 and a MSE of 0.1310.

Cite

CITATION STYLE

APA

Al Ghozali, I. H., Fathin, M. A., & Handoko, A. R. (2025). A Comparative Study of Machine Learning with Statistical Feature Selection for Risk Detection of Diabetic. Jurnal Ilmiah FIFO, 17(2), 102. https://doi.org/10.22441/fifo.2025.v17i2.001

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free