Downloads provided by UsageCounts
arXiv: 2207.00551
handle: 11584/432664
Neural networks are ubiquitous in applied machine learning for education. Their pervasive success in predictive performance comes alongside a severe weakness, the lack of explainability of their decisions, especially relevant in human-centric fields. We implement five state-of-the-art methodologies for explaining black-box machine learning models (LIME, PermutationSHAP, KernelSHAP, DiCE, CEM) and examine the strengths of each approach on the downstream task of student performance prediction for five massive open online courses. Our experiments demonstrate that the families of explainers do not agree with each other on feature importance for the same Bidirectional LSTM models with the same representative set of students. We use Principal Component Analysis, Jensen-Shannon distance, and Spearman's rank-order correlation to quantitatively cross-examine explanations across methods and courses. Furthermore, we validate explainer performance across curriculum-based prerequisite relationships. Our results come to the concerning conclusion that the choice of explainer is an important decision and is in fact paramount to the interpretation of the predictive results, even more so than the course the model is trained on. Source code and models are released at http://github.com/epfl-ml4ed/evaluating-explainers.
Accepted as a full paper at EDM 2022: The 15th International Conference on Educational Data Mining, 24-27 of July 2022, Durham
FOS: Computer and information sciences, Computer Science - Machine Learning, DiCE, MOOCs, CEM, LIME, LSTMs, Machine Learning (cs.LG), Computer Science - Computers and Society, SHAP, Explainable AI, Computers and Society (cs.CY), Student Performance Prediction, Counterfactuals
FOS: Computer and information sciences, Computer Science - Machine Learning, DiCE, MOOCs, CEM, LIME, LSTMs, Machine Learning (cs.LG), Computer Science - Computers and Society, SHAP, Explainable AI, Computers and Society (cs.CY), Student Performance Prediction, Counterfactuals
| selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | 7 | |
| popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network. | Top 10% | |
| influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | Average | |
| impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network. | Top 10% |
| views | 19 | |
| downloads | 17 |

Views provided by UsageCounts
Downloads provided by UsageCounts