Saved in:
| Main Authors: | Zhang, Emerald, Weaver, Julian, Santacruz, Samantha R, Castillo, Edward |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2509.23585 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AttnLRP: Attention-Aware Layer-Wise Relevance Propagation for Transformers
by: Achtibat, Reduan, et al.
Published: (2024)
by: Achtibat, Reduan, et al.
Published: (2024)
An accuracy-aware extension to LRP-based pruning for CNNs to prevent cascading accuracy degradation in data-scarce transfer learning
by: Yasui, Daisuke, et al.
Published: (2025)
by: Yasui, Daisuke, et al.
Published: (2025)
Value bounds and Convergence Analysis for Averages of LRP attributions
by: Binder, Alexander, et al.
Published: (2025)
by: Binder, Alexander, et al.
Published: (2025)
Unsupervised Interpretable Basis Extraction for Concept-Based Visual Explanations
by: Doumanoglou, Alexandros, et al.
Published: (2023)
by: Doumanoglou, Alexandros, et al.
Published: (2023)
Provably Better Explanations with Optimized Aggregation of Feature Attributions
by: Decker, Thomas, et al.
Published: (2024)
by: Decker, Thomas, et al.
Published: (2024)
Explainability of Point Cloud Neural Networks Using SMILE: Statistical Model-Agnostic Interpretability with Local Explanations
by: Ahmadi, Seyed Mohammad, et al.
Published: (2024)
by: Ahmadi, Seyed Mohammad, et al.
Published: (2024)
Explanation Bottleneck Models
by: Yamaguchi, Shin'ya, et al.
Published: (2024)
by: Yamaguchi, Shin'ya, et al.
Published: (2024)
Studying How to Efficiently and Effectively Guide Models with Explanations
by: Rao, Sukrut, et al.
Published: (2023)
by: Rao, Sukrut, et al.
Published: (2023)
Towards a Mechanistic Explanation of Diffusion Model Generalization
by: Niedoba, Matthew, et al.
Published: (2024)
by: Niedoba, Matthew, et al.
Published: (2024)
MEGL: Multimodal Explanation-Guided Learning
by: Zhang, Yifei, et al.
Published: (2024)
by: Zhang, Yifei, et al.
Published: (2024)
LINE: LLM-based Iterative Neuron Explanations for Vision Models
by: Zaigrajew, Vladimir, et al.
Published: (2026)
by: Zaigrajew, Vladimir, et al.
Published: (2026)
Leveraging Local Structure for Improving Model Explanations: An Information Propagation Approach
by: Yang, Ruo, et al.
Published: (2024)
by: Yang, Ruo, et al.
Published: (2024)
Probabilistic Conceptual Explainers: Trustworthy Conceptual Explanations for Vision Foundation Models
by: Wang, Hengyi, et al.
Published: (2024)
by: Wang, Hengyi, et al.
Published: (2024)
Technical Note: Defining and Quantifying AND-OR Interactions for Faithful and Concise Explanation of DNNs
by: Li, Mingjie, et al.
Published: (2023)
by: Li, Mingjie, et al.
Published: (2023)
Towards Multi-dimensional Explanation Alignment for Medical Classification
by: Hu, Lijie, et al.
Published: (2024)
by: Hu, Lijie, et al.
Published: (2024)
Statistically Significant Concept-based Explanation of Image Classifiers via Model Knockoffs
by: Xu, Kaiwen, et al.
Published: (2023)
by: Xu, Kaiwen, et al.
Published: (2023)
Neuron Abandoning Attention Flow: Visual Explanation of Dynamics inside CNN Models
by: Liao, Yi, et al.
Published: (2024)
by: Liao, Yi, et al.
Published: (2024)
Guaranteed Optimal Compositional Explanations for Neurons
by: La Rosa, Biagio, et al.
Published: (2025)
by: La Rosa, Biagio, et al.
Published: (2025)
Patronus: Interpretable Diffusion Models with Prototypes
by: Weng, Nina, et al.
Published: (2025)
by: Weng, Nina, et al.
Published: (2025)
Prisma: An Open Source Toolkit for Mechanistic Interpretability in Vision and Video
by: Joseph, Sonia, et al.
Published: (2025)
by: Joseph, Sonia, et al.
Published: (2025)
Scaling Image and Video Generation via Test-Time Evolutionary Search
by: He, Haoran, et al.
Published: (2025)
by: He, Haoran, et al.
Published: (2025)
On Spectral Properties of Gradient-based Explanation Methods
by: Mehrpanah, Amir, et al.
Published: (2025)
by: Mehrpanah, Amir, et al.
Published: (2025)
Measuring the (Un)Faithfulness of Concept-Based Explanations
by: Kumar, Shubham, et al.
Published: (2025)
by: Kumar, Shubham, et al.
Published: (2025)
Open Vocabulary Compositional Explanations for Neuron Alignment
by: La Rosa, Biagio, et al.
Published: (2025)
by: La Rosa, Biagio, et al.
Published: (2025)
Generative Emotion Cause Explanation in Multimodal Conversations
by: Wang, Lin, et al.
Published: (2024)
by: Wang, Lin, et al.
Published: (2024)
ViGText: Deepfake Image Detection with Vision-Language Model Explanations and Graph Neural Networks
by: ALBarqawi, Ahmad, et al.
Published: (2025)
by: ALBarqawi, Ahmad, et al.
Published: (2025)
Pixel-level Certified Explanations via Randomized Smoothing
by: Anani, Alaa, et al.
Published: (2025)
by: Anani, Alaa, et al.
Published: (2025)
On the Complexity-Faithfulness Trade-off of Gradient-Based Explanations
by: Mehrpanah, Amir, et al.
Published: (2025)
by: Mehrpanah, Amir, et al.
Published: (2025)
Relevant Irrelevance: Generating Alterfactual Explanations for Image Classifiers
by: Mertes, Silvan, et al.
Published: (2024)
by: Mertes, Silvan, et al.
Published: (2024)
Good Teachers Explain: Explanation-Enhanced Knowledge Distillation
by: Parchami-Araghi, Amin, et al.
Published: (2024)
by: Parchami-Araghi, Amin, et al.
Published: (2024)
DEAL: Disentangle and Localize Concept-level Explanations for VLMs
by: Li, Tang, et al.
Published: (2024)
by: Li, Tang, et al.
Published: (2024)
A Geometric Explanation of the Likelihood OOD Detection Paradox
by: Kamkari, Hamidreza, et al.
Published: (2024)
by: Kamkari, Hamidreza, et al.
Published: (2024)
Residualized Temporal Sparse Autoencoders for Interpreting Diffusion Models
by: Yeung, Calvin, et al.
Published: (2026)
by: Yeung, Calvin, et al.
Published: (2026)
Disentangled Explanations of Neural Network Predictions by Finding Relevant Subspaces
by: Chormai, Pattarawat, et al.
Published: (2022)
by: Chormai, Pattarawat, et al.
Published: (2022)
What Helps---and What Hurts: Bidirectional Explanations for Vision Transformers
by: Su, Qin, et al.
Published: (2026)
by: Su, Qin, et al.
Published: (2026)
BEE: Metric-Adapted Explanations via Baseline Exploration-Exploitation
by: Barkan, Oren, et al.
Published: (2024)
by: Barkan, Oren, et al.
Published: (2024)
Restyling Unsupervised Concept Based Interpretable Networks with Generative Models
by: Parekh, Jayneel, et al.
Published: (2024)
by: Parekh, Jayneel, et al.
Published: (2024)
Benchmarking the Influence of Pre-training on Explanation Performance in MR Image Classification
by: Oliveira, Marta, et al.
Published: (2023)
by: Oliveira, Marta, et al.
Published: (2023)
Walking the Web of Concept-Class Relationships in Incrementally Trained Interpretable Models
by: Agrawal, Susmit, et al.
Published: (2025)
by: Agrawal, Susmit, et al.
Published: (2025)
This Looks Better than That: Better Interpretable Models with ProtoPNeXt
by: Willard, Frank, et al.
Published: (2024)
by: Willard, Frank, et al.
Published: (2024)
Similar Items
-
AttnLRP: Attention-Aware Layer-Wise Relevance Propagation for Transformers
by: Achtibat, Reduan, et al.
Published: (2024) -
An accuracy-aware extension to LRP-based pruning for CNNs to prevent cascading accuracy degradation in data-scarce transfer learning
by: Yasui, Daisuke, et al.
Published: (2025) -
Value bounds and Convergence Analysis for Averages of LRP attributions
by: Binder, Alexander, et al.
Published: (2025) -
Unsupervised Interpretable Basis Extraction for Concept-Based Visual Explanations
by: Doumanoglou, Alexandros, et al.
Published: (2023) -
Provably Better Explanations with Optimized Aggregation of Feature Attributions
by: Decker, Thomas, et al.
Published: (2024)