Abstract
Existing methods for explaining black box learning models often focus on building local explanations of the models’ behaviour for particular data items. It is possible to create global explanations for all data items, but these explanations generally have low fidelity for complex black box models. We propose a new supervised manifold visualisation method, slisemap, that simultaneously finds local explanations for all data items and builds a (typically) two-dimensional global visualisation of the black box model such that data items with similar local explanations are projected nearby. We provide a mathematical derivation of our problem and an open source implementation implemented using the GPU-optimised PyTorch library. We compare slisemap to multiple popular dimensionality reduction methods and find that slisemap is able to utilise labelled data to create embeddings with consistent local white box models. We also compare slisemap to other model-agnostic local explanation methods and show that slisemap provides comparable explanations and that the visualisations can give a broader understanding of black box regression and classification models.
Author supplied keywords
Cite
CITATION STYLE
Björklund, A., Mäkelä, J., & Puolamäki, K. (2023). SLISEMAP: supervised dimensionality reduction through local explanations. Machine Learning, 112(1), 1–43. https://doi.org/10.1007/s10994-022-06261-1
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.