Exploring Predictive Insights on Student Success Using Explainable Machine Learning: A Synthetic Data Study

2Citations
Citations of this article
38Readers
Mendeley users who have this article in their library.

Abstract

Student success is a multifaceted outcome influenced by academic, behavioral, contextual, and socio-environmental factors. With the growing availability of educational data, machine learning (ML) offers promising tools to model complex, nonlinear relationships that go beyond traditional statistical methods. However, the lack of interpretability in many ML models remains a major obstacle for practical adoption in educational contexts. In this study, we apply explainable artificial intelligence (XAI) techniques—specifically SHAP (SHapley Additive exPlanations)—to analyze a synthetic dataset simulating diverse student profiles. Using LightGBM, we identify variables such as hours studied, attendance, and parental involvement as influential in predicting exam performance. While the results are not generalizable due to the artificial nature of the data, this study reframes its purpose as a methodological exploration rather than a claim of real-world actionable insights. Our findings demonstrate how interpretable ML can be used to build transparent analytic pipelines in education, setting the stage for future research using empirical datasets and real student data.

Cite

CITATION STYLE

APA

Santana-Perera, B., García-Barceló, C., González Arcas, M., & Gil, D. (2025). Exploring Predictive Insights on Student Success Using Explainable Machine Learning: A Synthetic Data Study. Information (Switzerland), 16(9). https://doi.org/10.3390/info16090763

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free