Navigating the Maze of Explainable AI: A Systematic Approach to Evaluating Methods and Metrics
Fuente:
arXiv
Saved in:
| Main Authors: | Klein, Lukas, Lüth, Carsten T., Schlegel, Udo, Bungert, Till J., El-Assady, Mennatallah, Jäger, Paul F. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Finally Outshining the Random Baseline: A Simple and Effective Solution for Active Learning in 3D Biomedical Imaging
by: Lüth, Carsten T., et al.
Published: (2026)
by: Lüth, Carsten T., et al.
Published: (2026)
Why context matters in VQA and Reasoning: Semantic interventions for VLM input modalities
by: Amara, Kenza, et al.
Published: (2024)
by: Amara, Kenza, et al.
Published: (2024)
nnActive: A Framework for Evaluation of Active Learning in 3D Biomedical Segmentation
by: Lüth, Carsten T., et al.
Published: (2025)
by: Lüth, Carsten T., et al.
Published: (2025)
Overcoming Common Flaws in the Evaluation of Selective Classification Systems
by: Traub, Jeremias, et al.
Published: (2024)
by: Traub, Jeremias, et al.
Published: (2024)
ValUES: A Framework for Systematic Validation of Uncertainty Estimation in Semantic Segmentation
by: Kahl, Kim-Celine, et al.
Published: (2024)
by: Kahl, Kim-Celine, et al.
Published: (2024)
SyntaxShap: Syntax-aware Explainability Method for Text Generation
by: Amara, Kenza, et al.
Published: (2024)
by: Amara, Kenza, et al.
Published: (2024)
SURE-VQA: Systematic Understanding of Robustness Evaluation in Medical VQA Tasks
by: Kahl, Kim-Celine, et al.
Published: (2024)
by: Kahl, Kim-Celine, et al.
Published: (2024)
A Meaningful Perturbation Metric for Evaluating Explainability Methods
by: Cohen, Danielle, et al.
Published: (2025)
by: Cohen, Danielle, et al.
Published: (2025)
Deconstructing Human-AI Collaboration: Agency, Interaction, and Adaptation
by: Holter, Steffen, et al.
Published: (2024)
by: Holter, Steffen, et al.
Published: (2024)
Deconstructing Human‐AI Collaboration: Agency, Interaction, and Adaptation
by: Steffen Holter, et al.
Published: (2024)
by: Steffen Holter, et al.
Published: (2024)
Challenges and Opportunities in Text Generation Explainability
by: Amara, Kenza, et al.
Published: (2024)
by: Amara, Kenza, et al.
Published: (2024)
Concept-Level Explainability for Auditing & Steering LLM Responses
by: Amara, Kenza, et al.
Published: (2025)
by: Amara, Kenza, et al.
Published: (2025)
On the Effectiveness of Methods and Metrics for Explainable AI in Remote Sensing Image Scene Classification
by: Klotz, Jonas, et al.
Published: (2025)
by: Klotz, Jonas, et al.
Published: (2025)
The Labyrinth of Links: Navigating the Associative Maze of Multi-modal LLMs
by: Li, Hong, et al.
Published: (2024)
by: Li, Hong, et al.
Published: (2024)
Towards Quantitative Evaluation of Explainable AI Methods for Deepfake Detection
by: Tsigos, Konstantinos, et al.
Published: (2024)
by: Tsigos, Konstantinos, et al.
Published: (2024)
The Weighting Game: Evaluating Quality of Explainability Methods
by: Raatikainen, Lassi, et al.
Published: (2022)
by: Raatikainen, Lassi, et al.
Published: (2022)
Evaluating Explainable AI Methods in Deep Learning Models for Early Detection of Cerebral Palsy
by: Pellano, Kimji N., et al.
Published: (2024)
by: Pellano, Kimji N., et al.
Published: (2024)
What Makes a Maze Look Like a Maze?
by: Hsu, Joy, et al.
Published: (2024)
by: Hsu, Joy, et al.
Published: (2024)
ODExAI: A Comprehensive Object Detection Explainable AI Evaluation
by: Nguyen, Loc Phuc Truong, et al.
Published: (2025)
by: Nguyen, Loc Phuc Truong, et al.
Published: (2025)
Metric for Evaluating Performance of Reference-Free Demorphing Methods
by: Shukla, Nitish, et al.
Published: (2025)
by: Shukla, Nitish, et al.
Published: (2025)
EXPERT: An Explainable Image Captioning Evaluation Metric with Structured Explanations
by: Kim, Hyunjong, et al.
Published: (2025)
by: Kim, Hyunjong, et al.
Published: (2025)
MetricNet: Recovering Metric Scale in Generative Navigation Policies
by: Nayak, Abhijeet, et al.
Published: (2025)
by: Nayak, Abhijeet, et al.
Published: (2025)
A Quantitative Evaluation Framework for Explainable AI in Semantic Segmentation
by: Hammoud, Reem, et al.
Published: (2025)
by: Hammoud, Reem, et al.
Published: (2025)
VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation
by: Ku, Max, et al.
Published: (2023)
by: Ku, Max, et al.
Published: (2023)
Video Models Reason Early: Exploiting Plan Commitment for Maze Solving
by: Newman, Kaleb, et al.
Published: (2026)
by: Newman, Kaleb, et al.
Published: (2026)
SplineFormer: An Explainable Transformer-Based Approach for Autonomous Endovascular Navigation
by: Jianu, Tudor, et al.
Published: (2025)
by: Jianu, Tudor, et al.
Published: (2025)
iNNspector: Visual, Interactive Deep Model Debugging
by: Spinner, Thilo, et al.
Published: (2024)
by: Spinner, Thilo, et al.
Published: (2024)
Challenges, Advances, and Evaluation Metrics in Medical Image Enhancement: A Systematic Literature Review
by: Chin, Chun Wai, et al.
Published: (2025)
by: Chin, Chun Wai, et al.
Published: (2025)
Explainable Metric Learning for Deflating Data Bias
by: Andrews, Emma, et al.
Published: (2024)
by: Andrews, Emma, et al.
Published: (2024)
Exploring Convolutional Neural Networks for Rice Grain Classification: An Explainable AI Approach
by: Asif, Muhammad Junaid, et al.
Published: (2025)
by: Asif, Muhammad Junaid, et al.
Published: (2025)
A Deep Learning Approach for Automated Skin Lesion Diagnosis with Explainable AI
by: Haque, Md. Maksudul, et al.
Published: (2026)
by: Haque, Md. Maksudul, et al.
Published: (2026)
Reasoning via Video: The First Evaluation of Video Models' Reasoning Abilities through Maze-Solving Tasks
by: Yang, Cheng, et al.
Published: (2025)
by: Yang, Cheng, et al.
Published: (2025)
Comparative Benchmarking of Failure Detection Methods in Medical Image Segmentation: Unveiling the Role of Confidence Aggregation
by: Zenk, Maximilian, et al.
Published: (2024)
by: Zenk, Maximilian, et al.
Published: (2024)
Explainable Convolutional Networks for Crater Detection and Lunar Landing Navigation
by: Song, Jianing, et al.
Published: (2024)
by: Song, Jianing, et al.
Published: (2024)
FunnyNodules: A Customizable Medical Dataset Tailored for Evaluating Explainable AI
by: Gallée, Luisa, et al.
Published: (2025)
by: Gallée, Luisa, et al.
Published: (2025)
Back to the Baseline: Examining Baseline Effects on Explainability Metrics
by: Picard, Agustin Martin, et al.
Published: (2025)
by: Picard, Agustin Martin, et al.
Published: (2025)
EPSM: A Novel Metric to Evaluate the Safety of Environmental Perception in Autonomous Driving
by: Gamerdinger, Jörg, et al.
Published: (2025)
by: Gamerdinger, Jörg, et al.
Published: (2025)
Unveiling the "Fairness Seesaw": Discovering and Mitigating Gender and Race Bias in Vision-Language Models
by: Lan, Jian, et al.
Published: (2025)
by: Lan, Jian, et al.
Published: (2025)
XIMAGENET-12: An Explainable AI Benchmark Dataset for Model Robustness Evaluation
by: Li, Qiang, et al.
Published: (2023)
by: Li, Qiang, et al.
Published: (2023)
Human Uncertainty-Aware Data Selection and Automatic Labeling in Visual Question Answering
by: Lan, Jian, et al.
Published: (2025)
by: Lan, Jian, et al.
Published: (2025)
Similar Items
-
Finally Outshining the Random Baseline: A Simple and Effective Solution for Active Learning in 3D Biomedical Imaging
by: Lüth, Carsten T., et al.
Published: (2026) -
Why context matters in VQA and Reasoning: Semantic interventions for VLM input modalities
by: Amara, Kenza, et al.
Published: (2024) -
nnActive: A Framework for Evaluation of Active Learning in 3D Biomedical Segmentation
by: Lüth, Carsten T., et al.
Published: (2025) -
Overcoming Common Flaws in the Evaluation of Selective Classification Systems
by: Traub, Jeremias, et al.
Published: (2024) -
ValUES: A Framework for Systematic Validation of Uncertainty Estimation in Semantic Segmentation
by: Kahl, Kim-Celine, et al.
Published: (2024)