Structural Compactness as a Complementary Criterion for Explanation Quality
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mesgari, Mohammad Mahdi, Ma, Jackie, Samek, Wojciech, Lapuschkin, Sebastian, Weber, Leander |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Relevance-driven Input Dropout: an Explanation-guided Regularization Technique
von: Gururaj, Shreyas, et al.
Veröffentlicht: (2025)
von: Gururaj, Shreyas, et al.
Veröffentlicht: (2025)
Understanding the (Extra-)Ordinary: Validating Deep Model Decisions with Prototypical Concept-based Explanations
von: Dreyer, Maximilian, et al.
Veröffentlicht: (2023)
von: Dreyer, Maximilian, et al.
Veröffentlicht: (2023)
Iterative Inference in a Chess-Playing Neural Network
von: Sandmann, Elias, et al.
Veröffentlicht: (2025)
von: Sandmann, Elias, et al.
Veröffentlicht: (2025)
See What I Mean? CUE: A Cognitive Model of Understanding Explanations
von: Labarta, Tobias, et al.
Veröffentlicht: (2025)
von: Labarta, Tobias, et al.
Veröffentlicht: (2025)
Explaining Predictive Uncertainty by Exposing Second-Order Effects
von: Bley, Florian, et al.
Veröffentlicht: (2024)
von: Bley, Florian, et al.
Veröffentlicht: (2024)
Efficient and Flexible Neural Network Training through Layer-wise Feedback Propagation
von: Weber, Leander, et al.
Veröffentlicht: (2023)
von: Weber, Leander, et al.
Veröffentlicht: (2023)
From Attribution Maps to Human-Understandable Explanations through Concept Relevance Propagation
von: Achtibat, Reduan, et al.
Veröffentlicht: (2022)
von: Achtibat, Reduan, et al.
Veröffentlicht: (2022)
Atlas-Alignment: Making Interpretability Transferable Across Language Models
von: Puri, Bruno, et al.
Veröffentlicht: (2025)
von: Puri, Bruno, et al.
Veröffentlicht: (2025)
Post-Hoc Concept Disentanglement: From Correlated to Isolated Concept Representations
von: Erogullari, Eren, et al.
Veröffentlicht: (2025)
von: Erogullari, Eren, et al.
Veröffentlicht: (2025)
Ensuring Medical AI Safety: Interpretability-Driven Detection and Mitigation of Spurious Model Behavior and Associated Data
von: Pahde, Frederik, et al.
Veröffentlicht: (2025)
von: Pahde, Frederik, et al.
Veröffentlicht: (2025)
Navigating Neural Space: Revisiting Concept Activation Vectors to Overcome Directional Divergence
von: Pahde, Frederik, et al.
Veröffentlicht: (2022)
von: Pahde, Frederik, et al.
Veröffentlicht: (2022)
X-SYS: A Reference Architecture for Interactive Explanation Systems
von: Labarta, Tobias, et al.
Veröffentlicht: (2026)
von: Labarta, Tobias, et al.
Veröffentlicht: (2026)
ECQ$^{\text{x}}$: Explainability-Driven Quantization for Low-Bit and Sparse DNNs
von: Becking, Daniel, et al.
Veröffentlicht: (2021)
von: Becking, Daniel, et al.
Veröffentlicht: (2021)
Human-Centered Evaluation of XAI Methods
von: Dawoud, Karam, et al.
Veröffentlicht: (2023)
von: Dawoud, Karam, et al.
Veröffentlicht: (2023)
Sparse, Efficient and Explainable Data Attribution with DualXDA
von: Yolcu, Galip Ümit, et al.
Veröffentlicht: (2024)
von: Yolcu, Galip Ümit, et al.
Veröffentlicht: (2024)
PURE: Turning Polysemantic Neurons Into Pure Features by Identifying Relevant Circuits
von: Dreyer, Maximilian, et al.
Veröffentlicht: (2024)
von: Dreyer, Maximilian, et al.
Veröffentlicht: (2024)
Reactive Model Correction: Mitigating Harm to Task-Relevant Features via Conditional Bias Suppression
von: Bareeva, Dilyara, et al.
Veröffentlicht: (2024)
von: Bareeva, Dilyara, et al.
Veröffentlicht: (2024)
A Fresh Look at Sanity Checks for Saliency Maps
von: Hedström, Anna, et al.
Veröffentlicht: (2024)
von: Hedström, Anna, et al.
Veröffentlicht: (2024)
Sanity Checks Revisited: An Exploration to Repair the Model Parameter Randomisation Test
von: Hedström, Anna, et al.
Veröffentlicht: (2024)
von: Hedström, Anna, et al.
Veröffentlicht: (2024)
From What to How: Attributing CLIP's Latent Components Reveals Unexpected Semantic Reliance
von: Dreyer, Maximilian, et al.
Veröffentlicht: (2025)
von: Dreyer, Maximilian, et al.
Veröffentlicht: (2025)
Improved discrete particle swarm optimization using Bee Algorithm and multi-parent crossover method (Case study: Allocation problem and benchmark functions)
von: Zibaei, Hamed, et al.
Veröffentlicht: (2024)
von: Zibaei, Hamed, et al.
Veröffentlicht: (2024)
Dyslexify: A Mechanistic Defense Against Typographic Attacks in CLIP
von: Hufe, Lorenz, et al.
Veröffentlicht: (2025)
von: Hufe, Lorenz, et al.
Veröffentlicht: (2025)
Circuit Insights: Towards Interpretability Beyond Activations
von: Golimblevskaia, Elena, et al.
Veröffentlicht: (2025)
von: Golimblevskaia, Elena, et al.
Veröffentlicht: (2025)
Synthetic Datasets for Machine Learning on Spatio-Temporal Graphs using PDEs
von: Arndt, Jost, et al.
Veröffentlicht: (2025)
von: Arndt, Jost, et al.
Veröffentlicht: (2025)
Pruning By Explaining Revisited: Optimizing Attribution Methods to Prune CNNs and Transformers
von: Hatefi, Sayed Mohammad Vakilzadeh, et al.
Veröffentlicht: (2024)
von: Hatefi, Sayed Mohammad Vakilzadeh, et al.
Veröffentlicht: (2024)
Mechanistic understanding and validation of large AI models with SemanticLens
von: Dreyer, Maximilian, et al.
Veröffentlicht: (2025)
von: Dreyer, Maximilian, et al.
Veröffentlicht: (2025)
Building Trust in PINNs: Error Estimation through Finite Difference Methods
von: Krasowski, Aleksander, et al.
Veröffentlicht: (2026)
von: Krasowski, Aleksander, et al.
Veröffentlicht: (2026)
Quanda: An Interpretability Toolkit for Training Data Attribution Evaluation and Beyond
von: Bareeva, Dilyara, et al.
Veröffentlicht: (2024)
von: Bareeva, Dilyara, et al.
Veröffentlicht: (2024)
FADE: Why Bad Descriptions Happen to Good Features
von: Puri, Bruno, et al.
Veröffentlicht: (2025)
von: Puri, Bruno, et al.
Veröffentlicht: (2025)
From Attribution to Action: A Human-Centered Application of Activation Steering
von: Labarta, Tobias, et al.
Veröffentlicht: (2026)
von: Labarta, Tobias, et al.
Veröffentlicht: (2026)
Attribution-Guided Pruning for Insight and Control: Circuit Discovery and Targeted Correction in Small-scale LLMs
von: Hatefi, Sayed Mohammad Vakilzadeh, et al.
Veröffentlicht: (2025)
von: Hatefi, Sayed Mohammad Vakilzadeh, et al.
Veröffentlicht: (2025)
AttnLRP: Attention-Aware Layer-Wise Relevance Propagation for Transformers
von: Achtibat, Reduan, et al.
Veröffentlicht: (2024)
von: Achtibat, Reduan, et al.
Veröffentlicht: (2024)
Model Science: getting serious about verification, explanation and control of AI systems
von: Biecek, Przemyslaw, et al.
Veröffentlicht: (2025)
von: Biecek, Przemyslaw, et al.
Veröffentlicht: (2025)
LieSolver: A PDE-constrained solver for IBVPs using Lie symmetries
von: Klausen, René P., et al.
Veröffentlicht: (2025)
von: Klausen, René P., et al.
Veröffentlicht: (2025)
Position: Explain to Question not to Justify
von: Biecek, Przemyslaw, et al.
Veröffentlicht: (2024)
von: Biecek, Przemyslaw, et al.
Veröffentlicht: (2024)
PINNfluence: Influence Functions for Physics-Informed Neural Networks
von: Naujoks, Jonas R., et al.
Veröffentlicht: (2024)
von: Naujoks, Jonas R., et al.
Veröffentlicht: (2024)
Deep Learning-based Multi Project InP Wafer Simulation for Unsupervised Surface Defect Detection
von: Cantú, Emílio Dolgener, et al.
Veröffentlicht: (2025)
von: Cantú, Emílio Dolgener, et al.
Veröffentlicht: (2025)
Strategic Dishonesty Can Undermine AI Safety Evaluations of Frontier LLMs
von: Panfilov, Alexander, et al.
Veröffentlicht: (2025)
von: Panfilov, Alexander, et al.
Veröffentlicht: (2025)
Leveraging Influence Functions for Resampling Data in Physics-Informed Neural Networks
von: Naujoks, Jonas R., et al.
Veröffentlicht: (2025)
von: Naujoks, Jonas R., et al.
Veröffentlicht: (2025)
Synthetic Generation of Dermatoscopic Images with GAN and Closed-Form Factorization
von: Mekala, Rohan Reddy, et al.
Veröffentlicht: (2024)
von: Mekala, Rohan Reddy, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Relevance-driven Input Dropout: an Explanation-guided Regularization Technique
von: Gururaj, Shreyas, et al.
Veröffentlicht: (2025) -
Understanding the (Extra-)Ordinary: Validating Deep Model Decisions with Prototypical Concept-based Explanations
von: Dreyer, Maximilian, et al.
Veröffentlicht: (2023) -
Iterative Inference in a Chess-Playing Neural Network
von: Sandmann, Elias, et al.
Veröffentlicht: (2025) -
See What I Mean? CUE: A Cognitive Model of Understanding Explanations
von: Labarta, Tobias, et al.
Veröffentlicht: (2025) -
Explaining Predictive Uncertainty by Exposing Second-Order Effects
von: Bley, Florian, et al.
Veröffentlicht: (2024)