From Flexibility to Manipulation: The Slippery Slope of XAI Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Wickstrøm, Kristoffer, Höhne, Marina Marie-Claire, Hedström, Anna |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Finding the right XAI method -- A Guide for the Evaluation and Ranking of Explainable AI Methods in Climate Science
by: Bommer, Philine, et al.
Published: (2023)
by: Bommer, Philine, et al.
Published: (2023)
A Fresh Look at Sanity Checks for Saliency Maps
by: Hedström, Anna, et al.
Published: (2024)
by: Hedström, Anna, et al.
Published: (2024)
Sanity Checks Revisited: An Exploration to Repair the Model Parameter Randomisation Test
by: Hedström, Anna, et al.
Published: (2024)
by: Hedström, Anna, et al.
Published: (2024)
CoSy: Evaluating Textual Explanations of Neurons
by: Kopf, Laura, et al.
Published: (2024)
by: Kopf, Laura, et al.
Published: (2024)
REPEAT: Improving Uncertainty Estimation in Representation Learning Explainability
by: Wickstrøm, Kristoffer K., et al.
Published: (2024)
by: Wickstrøm, Kristoffer K., et al.
Published: (2024)
Capturing Polysemanticity with PRISM: A Multi-Concept Feature Description Framework
by: Kopf, Laura, et al.
Published: (2025)
by: Kopf, Laura, et al.
Published: (2025)
Benchmarking XAI Explanations with Human-Aligned Evaluations
by: Kazmierczak, Rémi, et al.
Published: (2024)
by: Kazmierczak, Rémi, et al.
Published: (2024)
Exploring SAIG Methods for an Objective Evaluation of XAI
by: Miró-Nicolau, Miquel, et al.
Published: (2026)
by: Miró-Nicolau, Miquel, et al.
Published: (2026)
Evaluate with the Inverse: Efficient Approximation of Latent Explanation Quality Distribution
by: Eiras-Franco, Carlos, et al.
Published: (2025)
by: Eiras-Franco, Carlos, et al.
Published: (2025)
Uncertainty Gating for Cost-Aware Explainable Artificial Intelligence
by: Mikriukov, Georgii, et al.
Published: (2026)
by: Mikriukov, Georgii, et al.
Published: (2026)
Manipulating Feature Visualizations with Gradient Slingshots
by: Bareeva, Dilyara, et al.
Published: (2024)
by: Bareeva, Dilyara, et al.
Published: (2024)
Quanda: An Interpretability Toolkit for Training Data Attribution Evaluation and Beyond
by: Bareeva, Dilyara, et al.
Published: (2024)
by: Bareeva, Dilyara, et al.
Published: (2024)
Forms of Understanding for XAI-Explanations
by: Buschmeier, Hendrik, et al.
Published: (2023)
by: Buschmeier, Hendrik, et al.
Published: (2023)
Prototypical Self-Explainable Models Without Re-training
by: Gautam, Srishti, et al.
Published: (2023)
by: Gautam, Srishti, et al.
Published: (2023)
Trustworthy XAI and Application
by: Nasim, MD Abdullah Al, et al.
Published: (2024)
by: Nasim, MD Abdullah Al, et al.
Published: (2024)
XEQ Scale for Evaluating XAI Experience Quality
by: Wijekoon, Anjana, et al.
Published: (2024)
by: Wijekoon, Anjana, et al.
Published: (2024)
To Steer or Not to Steer? Mechanistic Error Reduction with Abstention for Language Models
by: Hedström, Anna, et al.
Published: (2025)
by: Hedström, Anna, et al.
Published: (2025)
Labeling Neural Representations with Inverse Recognition
by: Bykov, Kirill, et al.
Published: (2023)
by: Bykov, Kirill, et al.
Published: (2023)
OpenXAI: Towards a Transparent Evaluation of Model Explanations
by: Agarwal, Chirag, et al.
Published: (2022)
by: Agarwal, Chirag, et al.
Published: (2022)
Human-Centered Evaluation of XAI Methods
by: Dawoud, Karam, et al.
Published: (2023)
by: Dawoud, Karam, et al.
Published: (2023)
Evaluation Cards for XAI Metrics
by: Gipiškis, Rokas, et al.
Published: (2026)
by: Gipiškis, Rokas, et al.
Published: (2026)
From Black Boxes to Conversations: Incorporating XAI in a Conversational Agent
by: Nguyen, Van Bach, et al.
Published: (2022)
by: Nguyen, Van Bach, et al.
Published: (2022)
Deep Learning Meets Teleconnections: Improving S2S Predictions for European Winter Weather
by: Bommer, Philine L., et al.
Published: (2025)
by: Bommer, Philine L., et al.
Published: (2025)
AI Readiness in Healthcare through Storytelling XAI
by: Dubey, Akshat, et al.
Published: (2024)
by: Dubey, Akshat, et al.
Published: (2024)
The Role of XAI in Transforming Aeronautics and Aerospace Systems
by: Zorita, Francisco Javier Cantero, et al.
Published: (2024)
by: Zorita, Francisco Javier Cantero, et al.
Published: (2024)
Hacking a surrogate model approach to XAI
by: Wilhelm, Alexander, et al.
Published: (2024)
by: Wilhelm, Alexander, et al.
Published: (2024)
Guidelines For The Choice Of The Baseline in XAI Attribution Methods
by: Morasso, Cristian, et al.
Published: (2025)
by: Morasso, Cristian, et al.
Published: (2025)
Evaluation of Black-Box XAI Approaches for Predictors of Values of Boolean Formulae
by: Armoni-Friedmann, Stav, et al.
Published: (2025)
by: Armoni-Friedmann, Stav, et al.
Published: (2025)
STEER: Flexible Robotic Manipulation via Dense Language Grounding
by: Smith, Laura, et al.
Published: (2024)
by: Smith, Laura, et al.
Published: (2024)
Evaluating Explainability in Safety-Critical ATR Systems: Limitations of Post-Hoc Methods and Paths Toward Robust XAI
by: Buhrmester, Vanessa, et al.
Published: (2026)
by: Buhrmester, Vanessa, et al.
Published: (2026)
UbiQTree: Uncertainty Quantification in XAI with Tree Ensembles
by: Dubey, Akshat, et al.
Published: (2025)
by: Dubey, Akshat, et al.
Published: (2025)
Benchmarking Instance-Centric Counterfactual Algorithms for XAI: From White Box to Black Box
by: Moreira, Catarina, et al.
Published: (2022)
by: Moreira, Catarina, et al.
Published: (2022)
Quantifying True Robustness: Synonymity-Weighted Similarity for Trustworthy XAI Evaluation
by: Burger, Christopher
Published: (2025)
by: Burger, Christopher
Published: (2025)
A Mechanistic Explanatory Strategy for XAI
by: Rabiza, Marcin
Published: (2024)
by: Rabiza, Marcin
Published: (2024)
Kantian-Utilitarian XAI: Meta-Explained
by: Atf, Zahra, et al.
Published: (2025)
by: Atf, Zahra, et al.
Published: (2025)
Understanding XAI Through the Philosopher's Lens: A Historical Perspective
by: Mattioli, Martina, et al.
Published: (2024)
by: Mattioli, Martina, et al.
Published: (2024)
Why do explanations fail? A typology and discussion on failures in XAI
by: Bove, Clara, et al.
Published: (2024)
by: Bove, Clara, et al.
Published: (2024)
Is Conversational XAI All You Need? Human-AI Decision Making With a Conversational XAI Assistant
by: He, Gaole, et al.
Published: (2025)
by: He, Gaole, et al.
Published: (2025)
From Questions to Insights: Exploring XAI Challenges Reported on Stack Overflow Questions
by: Roy, Saumendu, et al.
Published: (2025)
by: Roy, Saumendu, et al.
Published: (2025)
Evaluating the Effectiveness of XAI Techniques for Encoder-Based Language Models
by: Mersha, Melkamu Abay, et al.
Published: (2025)
by: Mersha, Melkamu Abay, et al.
Published: (2025)
Similar Items
-
Finding the right XAI method -- A Guide for the Evaluation and Ranking of Explainable AI Methods in Climate Science
by: Bommer, Philine, et al.
Published: (2023) -
A Fresh Look at Sanity Checks for Saliency Maps
by: Hedström, Anna, et al.
Published: (2024) -
Sanity Checks Revisited: An Exploration to Repair the Model Parameter Randomisation Test
by: Hedström, Anna, et al.
Published: (2024) -
CoSy: Evaluating Textual Explanations of Neurons
by: Kopf, Laura, et al.
Published: (2024) -
REPEAT: Improving Uncertainty Estimation in Representation Learning Explainability
by: Wickstrøm, Kristoffer K., et al.
Published: (2024)