Evaluation of post-hoc interpretability methods in time-series classification

Fuente: arXiv
Gespeichert in:
Bibliographische Detailangaben
Hauptverfasser: Turbé, Hugues, Bjelogrlic, Mina, Lovis, Christian, Mengaldo, Gianmarco
Format: Preprint
Veröffentlicht: 2022
Schlagworte:
Online-Zugang:
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
_version_ 1866912146793693184
author Turbé, Hugues
Bjelogrlic, Mina
Lovis, Christian
Mengaldo, Gianmarco
author_facet Turbé, Hugues
Bjelogrlic, Mina
Lovis, Christian
Mengaldo, Gianmarco
contents Post-hoc interpretability methods are critical tools to explain neural-network results. Several post-hoc methods have emerged in recent years, but when applied to a given task, they produce different results, raising the question of which method is the most suitable to provide correct post-hoc interpretability. To understand the performance of each method, quantitative evaluation of interpretability methods is essential. However, currently available frameworks have several drawbacks which hinders the adoption of post-hoc interpretability methods, especially in high-risk sectors. In this work, we propose a framework with quantitative metrics to assess the performance of existing post-hoc interpretability methods in particular in time series classification. We show that several drawbacks identified in the literature are addressed, namely dependence on human judgement, retraining, and shift in the data distribution when occluding samples. We additionally design a synthetic dataset with known discriminative features and tunable complexity. The proposed methodology and quantitative metrics can be used to understand the reliability of interpretability methods results obtained in practical applications. In turn, they can be embedded within operational workflows in critical fields that require accurate interpretability results for e.g., regulatory policies.
format Preprint
id arxiv_https___arxiv_org_abs_2202_05656
institution arXiv
publishDate 2022
record_format arxiv
spellingShingle Evaluation of post-hoc interpretability methods in time-series classification
Turbé, Hugues
Bjelogrlic, Mina
Lovis, Christian
Mengaldo, Gianmarco
Machine Learning
Artificial Intelligence
I.2.6
Post-hoc interpretability methods are critical tools to explain neural-network results. Several post-hoc methods have emerged in recent years, but when applied to a given task, they produce different results, raising the question of which method is the most suitable to provide correct post-hoc interpretability. To understand the performance of each method, quantitative evaluation of interpretability methods is essential. However, currently available frameworks have several drawbacks which hinders the adoption of post-hoc interpretability methods, especially in high-risk sectors. In this work, we propose a framework with quantitative metrics to assess the performance of existing post-hoc interpretability methods in particular in time series classification. We show that several drawbacks identified in the literature are addressed, namely dependence on human judgement, retraining, and shift in the data distribution when occluding samples. We additionally design a synthetic dataset with known discriminative features and tunable complexity. The proposed methodology and quantitative metrics can be used to understand the reliability of interpretability methods results obtained in practical applications. In turn, they can be embedded within operational workflows in critical fields that require accurate interpretability results for e.g., regulatory policies.
title Evaluation of post-hoc interpretability methods in time-series classification
topic Machine Learning
Artificial Intelligence
I.2.6
url https://arxiv.org/abs/2202.05656