
arXiv: 2307.10654
A very popular model-agnostic technique for explaining predictive models is the SHapley Additive exPlanation (SHAP). The two most popular versions of SHAP are a conditional expectation version and an unconditional expectation version (the latter is also known as interventional SHAP). Except for tree-based methods, usually the unconditional version is used (for computational reasons). We provide a (surrogate) neural network approach which allows us to efficiently calculate the conditional version for both neural networks and other regression models, and which properly considers the dependence structure in the feature components. This proposal is also useful to provide drop1 and anova analyses in complex regression models which are similar to their generalized linear model (GLM) counterparts, and we provide a partial dependence plot (PDP) counterpart that considers the right dependence structure in the feature components.
24 pages, 9 figures
FOS: Computer and information sciences, Computer Science - Machine Learning, I.2.6, I.6.4, G.3, Machine Learning (stat.ML), Statistics - Applications, I.6.4; I.2.6; G.3, Machine Learning (cs.LG), Computational Engineering, Finance, and Science (cs.CE), Statistics - Machine Learning, 62J10, 62J12, Applications (stat.AP), Computer Science - Computational Engineering, Finance, and Science
FOS: Computer and information sciences, Computer Science - Machine Learning, I.2.6, I.6.4, G.3, Machine Learning (stat.ML), Statistics - Applications, I.6.4; I.2.6; G.3, Machine Learning (cs.LG), Computational Engineering, Finance, and Science (cs.CE), Statistics - Machine Learning, 62J10, 62J12, Applications (stat.AP), Computer Science - Computational Engineering, Finance, and Science
| selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | 5 | |
| popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network. | Top 10% | |
| influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | Average | |
| impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network. | Top 10% |
