Do Not Trust a Model because It is Confident: Uncovering and Characterizing Unknown Unknowns to Student Success Predictors in Online-Based Learning

Roberta Galici; Tanja Kaser; Gianni Fenu; Mirko Marras

Conference Proceedings

Do Not Trust a Model because It is Confident: Uncovering and Characterizing Unknown Unknowns to Student Success Predictors in Online-Based Learning

ACM International Conference Proceeding Series (2023) 441-452

DOI: 10.1145/3576050.3576148

7Citations

19Readers

Get full text

Abstract

Student success models might be prone to develop weak spots, i.e., examples hard to accurately classify due to insufficient representation during model creation. This weakness is one of the main factors undermining users' trust, since model predictions could for instance lead an instructor to not intervene on a student in need. In this paper, we unveil the need of detecting and characterizing unknown unknowns in student success prediction in order to better understand when models may fail. Unknown unknowns include the students for which the model is highly confident in its predictions, but is actually wrong. Therefore, we cannot solely rely on the model's confidence when evaluating the predictions quality. We first introduce a framework for the identification and characterization of unknown unknowns. We then assess its informativeness on log data collected from flipped courses and online courses using quantitative analyses and interviews with instructors. Our results show that unknown unknowns are a critical issue in this domain and that our framework can be applied to support their detection. The source code is available at https://github.com/epfl-ml4ed/unknown-unknowns.

Author supplied keywords

Cite

CITATION STYLE

APA

Galici, R., Kaser, T., Fenu, G., & Marras, M. (2023). Do Not Trust a Model because It is Confident: Uncovering and Characterizing Unknown Unknowns to Student Success Predictors in Online-Based Learning. In ACM International Conference Proceeding Series (pp. 441–452). Association for Computing Machinery. https://doi.org/10.1145/3576050.3576148

Do Not Trust a Model because It is Confident: Uncovering and Characterizing Unknown Unknowns to Student Success Predictors in Online-Based Learning

Abstract

Author supplied keywords

Cite

Register to see more suggestions