Solving the enigma: Enhancing faithfulness and comprehensibility in explanations of deep networks

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Mamalakis, Michail, Mamalakis, Antonios, Agartz, Ingrid, Mørch-Johnsen, Lynn Egeland, Murray, Graham, Suckling, John, Lio, Pietro
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866915202315845632
author Mamalakis, Michail
Mamalakis, Antonios
Agartz, Ingrid
Mørch-Johnsen, Lynn Egeland
Murray, Graham
Suckling, John
Lio, Pietro
author_facet Mamalakis, Michail
Mamalakis, Antonios
Agartz, Ingrid
Mørch-Johnsen, Lynn Egeland
Murray, Graham
Suckling, John
Lio, Pietro
contents The accelerated progress of artificial intelligence (AI) has popularized deep learning models across various domains, yet their inherent opacity poses challenges, particularly in critical fields like healthcare, medicine, and the geosciences. Explainable AI (XAI) has emerged to shed light on these 'black box' models, aiding in deciphering their decision-making processes. However, different XAI methods often produce significantly different explanations, leading to high inter-method variability that increases uncertainty and undermines trust in deep networks' predictions. In this study, we address this challenge by introducing a novel framework designed to enhance the explainability of deep networks through a dual focus on maximizing both accuracy and comprehensibility in the explanations. Our framework integrates outputs from multiple established XAI methods and leverages a non-linear neural network model, termed the 'explanation optimizer,' to construct a unified, optimal explanation. The optimizer evaluates explanations using two key metrics: faithfulness (accuracy in reflecting the network's decisions) and complexity (comprehensibility). By balancing these, it provides accurate and accessible explanations, addressing a key XAI limitation. Experiments on multi-class and binary classification in 2D object and 3D neuroscience imaging confirm its efficacy. Our optimizer achieved faithfulness scores 155% and 63% higher than the best XAI methods in 3D and 2D tasks, respectively, while also reducing complexity for better understanding. These results demonstrate that optimal explanations based on specific quality criteria are achievable, offering a solution to the issue of inter-method variability in the current XAI literature and supporting more trustworthy deep network predictions
format Preprint
id arxiv_https___arxiv_org_abs_2405_10008
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Solving the enigma: Enhancing faithfulness and comprehensibility in explanations of deep networks
Mamalakis, Michail
Mamalakis, Antonios
Agartz, Ingrid
Mørch-Johnsen, Lynn Egeland
Murray, Graham
Suckling, John
Lio, Pietro
Computer Vision and Pattern Recognition
The accelerated progress of artificial intelligence (AI) has popularized deep learning models across various domains, yet their inherent opacity poses challenges, particularly in critical fields like healthcare, medicine, and the geosciences. Explainable AI (XAI) has emerged to shed light on these 'black box' models, aiding in deciphering their decision-making processes. However, different XAI methods often produce significantly different explanations, leading to high inter-method variability that increases uncertainty and undermines trust in deep networks' predictions. In this study, we address this challenge by introducing a novel framework designed to enhance the explainability of deep networks through a dual focus on maximizing both accuracy and comprehensibility in the explanations. Our framework integrates outputs from multiple established XAI methods and leverages a non-linear neural network model, termed the 'explanation optimizer,' to construct a unified, optimal explanation. The optimizer evaluates explanations using two key metrics: faithfulness (accuracy in reflecting the network's decisions) and complexity (comprehensibility). By balancing these, it provides accurate and accessible explanations, addressing a key XAI limitation. Experiments on multi-class and binary classification in 2D object and 3D neuroscience imaging confirm its efficacy. Our optimizer achieved faithfulness scores 155% and 63% higher than the best XAI methods in 3D and 2D tasks, respectively, while also reducing complexity for better understanding. These results demonstrate that optimal explanations based on specific quality criteria are achievable, offering a solution to the issue of inter-method variability in the current XAI literature and supporting more trustworthy deep network predictions
title Solving the enigma: Enhancing faithfulness and comprehensibility in explanations of deep networks
topic Computer Vision and Pattern Recognition
url https://arxiv.org/abs/2405.10008