How Reliable and Stable are Explanations of XAI Methods?
Fuente:
arXiv
Saved in:
| Main Authors: | Ribeiro, José, Cardoso, Lucas, Santos, Vitor, Carvalho, Eduardo, Carneiro, Níkolas, Alves, Ronnie |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Black Box Model Explanations and the Human Interpretability Expectations -- An Analysis in the Context of Homicide Prediction
by: Ribeiro, José, et al.
Published: (2022)
by: Ribeiro, José, et al.
Published: (2022)
Explanations Based on Item Response Theory (eXirt): A Model-Specific Method to Explain Tree-Ensemble Model in Trust Perspective
by: Ribeiro, José, et al.
Published: (2022)
by: Ribeiro, José, et al.
Published: (2022)
Enhancing Classifier Evaluation: A Fairer Benchmarking Strategy Based on Ability and Robustness
by: Cardoso, Lucas, et al.
Published: (2025)
by: Cardoso, Lucas, et al.
Published: (2025)
Beyond Random Sampling: Instance Quality-Based Data Partitioning via Item Response Theory
by: Cardoso, Lucas, et al.
Published: (2025)
by: Cardoso, Lucas, et al.
Published: (2025)
Standing on the shoulders of giants
by: Cardoso, Lucas Felipe Ferraro, et al.
Published: (2024)
by: Cardoso, Lucas Felipe Ferraro, et al.
Published: (2024)
Evaluating Model Explanations without Ground Truth
by: Rawal, Kaivalya, et al.
Published: (2025)
by: Rawal, Kaivalya, et al.
Published: (2025)
Deep Memory Search: A Metaheuristic Approach for Optimizing Heuristic Search
by: Hedar, Abdel-Rahman, et al.
Published: (2024)
by: Hedar, Abdel-Rahman, et al.
Published: (2024)
Relevance-driven Input Dropout: an Explanation-guided Regularization Technique
by: Gururaj, Shreyas, et al.
Published: (2025)
by: Gururaj, Shreyas, et al.
Published: (2025)
Logic-based Explanations for Linear Support Vector Classifiers with Reject Option
by: Filho, Francisco Mateus Rocha, et al.
Published: (2024)
by: Filho, Francisco Mateus Rocha, et al.
Published: (2024)
FORCE: Feature-Oriented Representation with Clustering and Explanation
by: Mukherjee, Rishav, et al.
Published: (2025)
by: Mukherjee, Rishav, et al.
Published: (2025)
Neural Reasoning Networks: Efficient Interpretable Neural Networks With Automatic Textual Explanations
by: Carrow, Stephen, et al.
Published: (2024)
by: Carrow, Stephen, et al.
Published: (2024)
Extending Differential Temporal Difference Methods for Episodic Problems
by: De Asis, Kris, et al.
Published: (2026)
by: De Asis, Kris, et al.
Published: (2026)
LLM-Driven Intrinsic Motivation for Sparse Reward Reinforcement Learning
by: Quadros, André, et al.
Published: (2025)
by: Quadros, André, et al.
Published: (2025)
The Geometry of Thought: How Scale Restructures Reasoning In Large Language Models
by: Anderson, Samuel Cyrenius
Published: (2026)
by: Anderson, Samuel Cyrenius
Published: (2026)
Energy Equity, Infrastructure and Demographic Analysis with XAI Methods
by: Shrestha, Sarahana, et al.
Published: (2025)
by: Shrestha, Sarahana, et al.
Published: (2025)
How Does Unfaithful Reasoning Emerge from Autoregressive Training? A Study of Synthetic Experiments
by: Wang, Fuxin, et al.
Published: (2026)
by: Wang, Fuxin, et al.
Published: (2026)
1 bit is all we need: binary normalized neural networks
by: Cabral, Eduardo Lobo Lustoda, et al.
Published: (2025)
by: Cabral, Eduardo Lobo Lustoda, et al.
Published: (2025)
First-Mover Bias in Gradient Boosting Explanations: Mechanism, Detection, and Resolution
by: Caraker, Drake, et al.
Published: (2026)
by: Caraker, Drake, et al.
Published: (2026)
XAI for Coding Agent Failures: Transforming Raw Execution Traces into Actionable Insights
by: Joshi, Arun
Published: (2026)
by: Joshi, Arun
Published: (2026)
Reliable Unlearning Harmful Information in LLMs with Metamorphosis Representation Projection
by: Wu, Chengcan, et al.
Published: (2025)
by: Wu, Chengcan, et al.
Published: (2025)
Diagnosing and Addressing Pitfalls in KG-RAG Datasets: Toward More Reliable Benchmarking
by: Zhang, Liangliang, et al.
Published: (2025)
by: Zhang, Liangliang, et al.
Published: (2025)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
by: Fadli, Samih
Published: (2025)
by: Fadli, Samih
Published: (2025)
Fusion-Based Neural Generalization for Predicting Temperature Fields in Industrial PET Preform Heating
by: Alsheikh, Ahmad, et al.
Published: (2025)
by: Alsheikh, Ahmad, et al.
Published: (2025)
Beyond Hallucinations: A Composite Score for Measuring Reliability in Open-Source Large Language Models
by: Salla, Rohit Kumar, et al.
Published: (2025)
by: Salla, Rohit Kumar, et al.
Published: (2025)
One-vs.-One Mitigation of Intersectional Bias: A General Method to Extend Fairness-Aware Binary Classification
by: Kobayashi, Kenji, et al.
Published: (2020)
by: Kobayashi, Kenji, et al.
Published: (2020)
DELTA: Variational Disentangled Learning for Privacy-Preserving Data Reprogramming
by: Malarkkan, Arun Vignesh, et al.
Published: (2025)
by: Malarkkan, Arun Vignesh, et al.
Published: (2025)
TS-Insight: Visualizing Thompson Sampling for Verification and XAI
by: Vares, Parsa, et al.
Published: (2025)
by: Vares, Parsa, et al.
Published: (2025)
NeSIG: A Neuro-Symbolic Method for Learning to Generate Planning Problems
by: Núñez-Molina, Carlos, et al.
Published: (2023)
by: Núñez-Molina, Carlos, et al.
Published: (2023)
TACO: Tackling Over-correction in Federated Learning with Tailored Adaptive Correction
by: Liu, Weijie, et al.
Published: (2025)
by: Liu, Weijie, et al.
Published: (2025)
DataRater: Meta-Learned Dataset Curation
by: Calian, Dan A., et al.
Published: (2025)
by: Calian, Dan A., et al.
Published: (2025)
Extrinsicaly Rewarded Soft Q Imitation Learning with Discriminator
by: Furuyama, Ryoma, et al.
Published: (2024)
by: Furuyama, Ryoma, et al.
Published: (2024)
Manipulating Predictions over Discrete Inputs in Machine Teaching
by: Wu, Xiaodong, et al.
Published: (2024)
by: Wu, Xiaodong, et al.
Published: (2024)
Axiomatic Characterisations of Sample-based Explainers
by: Amgoud, Leila, et al.
Published: (2024)
by: Amgoud, Leila, et al.
Published: (2024)
ParalESN: Enabling parallel information processing in Reservoir Computing
by: Pinna, Matteo, et al.
Published: (2026)
by: Pinna, Matteo, et al.
Published: (2026)
The Lattice Geometry of Neural Network Quantization -- A Short Equivalence Proof of GPTQ and Babai's Algorithm
by: Birnick, Johann
Published: (2025)
by: Birnick, Johann
Published: (2025)
Evaluation of post-hoc interpretability methods in time-series classification
by: Turbé, Hugues, et al.
Published: (2022)
by: Turbé, Hugues, et al.
Published: (2022)
Resilience to the Flowing Unknown: an Open Set Recognition Framework for Data Streams
by: Barcina-Blanco, Marcos, et al.
Published: (2024)
by: Barcina-Blanco, Marcos, et al.
Published: (2024)
Scaling Value Iteration Networks to 5000 Layers for Extreme Long-Term Planning
by: Wang, Yuhui, et al.
Published: (2024)
by: Wang, Yuhui, et al.
Published: (2024)
Understanding Goal Generalisation in Sequential Reinforcement Learning
by: Brown, Jason Ross, et al.
Published: (2026)
by: Brown, Jason Ross, et al.
Published: (2026)
Unlearning's Blind Spots: Over-Unlearning and Prototypical Relearning Attack
by: Ha, SeungBum, et al.
Published: (2025)
by: Ha, SeungBum, et al.
Published: (2025)
Similar Items
-
Black Box Model Explanations and the Human Interpretability Expectations -- An Analysis in the Context of Homicide Prediction
by: Ribeiro, José, et al.
Published: (2022) -
Explanations Based on Item Response Theory (eXirt): A Model-Specific Method to Explain Tree-Ensemble Model in Trust Perspective
by: Ribeiro, José, et al.
Published: (2022) -
Enhancing Classifier Evaluation: A Fairer Benchmarking Strategy Based on Ability and Robustness
by: Cardoso, Lucas, et al.
Published: (2025) -
Beyond Random Sampling: Instance Quality-Based Data Partitioning via Item Response Theory
by: Cardoso, Lucas, et al.
Published: (2025) -
Standing on the shoulders of giants
by: Cardoso, Lucas Felipe Ferraro, et al.
Published: (2024)