Predicting Student Dropout Risk in Online Learning Using Stacked Ensemble Machine Learning and Explainable AI Techniques

  • Ajayi O
N/ACitations
Citations of this article
13Readers
Mendeley users who have this article in their library.

Abstract

Predicting student dropout in online learning platforms such as MOOCs and institutional LMS platforms, is a critical challenge in educational data mining. Although numerous machine learning models have been proposed to predict dropout likelihood, the lack of model interpretability has limited their practical deployment in educational settings. This paper proposes a stacked ensemble machine learning model combining Logistic Regression, Random Forest, and XGBoost, with explainable AI techniques to identify at-risk learners using behavioral and demographic features. The dataset, obtained from Kaggle's MOOC Dropout Prediction challenge, was cleaned, balanced, and subjected to feature selection to prevent information leakage. With SHAP interpretability, the model achieves an accuracy of 65%, ROC AUC of 0.71, and PR AUC of 0.73. Our results show that dropout prediction is feasible using early behavioral data, and stacked models offer a promising balance of performance and transparency. This work contributes a replicable, explainable architecture suitable for real-time educational intervention systems.

Cite

CITATION STYLE

APA

Ajayi, O. O. (2025). Predicting Student Dropout Risk in Online Learning Using Stacked Ensemble Machine Learning and Explainable AI Techniques. International Journal of Computer Applications, 187(40), 26–29. https://doi.org/10.5120/ijca2025925707

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free