Saved in:
Bibliographic Details
Main Authors: Anwar, Saif, Griffiths, Nathan, Bhalerao, Abhir, Popham, Thomas
Format: Preprint
Published: 2024
Subjects:
Online Access:https://arxiv.org/abs/2408.10085
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866911994406240256
author Anwar, Saif
Griffiths, Nathan
Bhalerao, Abhir
Popham, Thomas
author_facet Anwar, Saif
Griffiths, Nathan
Bhalerao, Abhir
Popham, Thomas
contents Existing local Explainable AI (XAI) methods, such as LIME, select a region of the input space in the vicinity of a given input instance, for which they approximate the behaviour of a model using a simpler and more interpretable surrogate model. The size of this region is often controlled by a user-defined locality hyperparameter. In this paper, we demonstrate the difficulties associated with defining a suitable locality size to capture impactful model behaviour, as well as the inadequacy of using a single locality size to explain all predictions. We propose a novel method, MASALA, for generating explanations, which automatically determines the appropriate local region of impactful model behaviour for each individual instance being explained. MASALA approximates the local behaviour used by a complex model to make a prediction by fitting a linear surrogate model to a set of points which experience similar model behaviour. These points are found by clustering the input space into regions of linear behavioural trends exhibited by the model. We compare the fidelity and consistency of explanations generated by our method with existing local XAI methods, namely LIME and CHILLI. Experiments on the PHM08 and MIDAS datasets show that our method produces more faithful and consistent explanations than existing methods, without the need to define any sensitive locality hyperparameters.
format Preprint
id arxiv_https___arxiv_org_abs_2408_10085
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle MASALA: Model-Agnostic Surrogate Explanations by Locality Adaptation
Anwar, Saif
Griffiths, Nathan
Bhalerao, Abhir
Popham, Thomas
Machine Learning
Existing local Explainable AI (XAI) methods, such as LIME, select a region of the input space in the vicinity of a given input instance, for which they approximate the behaviour of a model using a simpler and more interpretable surrogate model. The size of this region is often controlled by a user-defined locality hyperparameter. In this paper, we demonstrate the difficulties associated with defining a suitable locality size to capture impactful model behaviour, as well as the inadequacy of using a single locality size to explain all predictions. We propose a novel method, MASALA, for generating explanations, which automatically determines the appropriate local region of impactful model behaviour for each individual instance being explained. MASALA approximates the local behaviour used by a complex model to make a prediction by fitting a linear surrogate model to a set of points which experience similar model behaviour. These points are found by clustering the input space into regions of linear behavioural trends exhibited by the model. We compare the fidelity and consistency of explanations generated by our method with existing local XAI methods, namely LIME and CHILLI. Experiments on the PHM08 and MIDAS datasets show that our method produces more faithful and consistent explanations than existing methods, without the need to define any sensitive locality hyperparameters.
title MASALA: Model-Agnostic Surrogate Explanations by Locality Adaptation
topic Machine Learning
url https://arxiv.org/abs/2408.10085