Measuring the (Un)Faithfulness of Concept-Based Explanations
Fuente:
arXiv
Saved in:
| Main Authors: | Kumar, Shubham, Ahuja, Narendra |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
On the Complexity-Faithfulness Trade-off of Gradient-Based Explanations
by: Mehrpanah, Amir, et al.
Published: (2025)
by: Mehrpanah, Amir, et al.
Published: (2025)
Piecewise-Linear Manifolds for Deep Metric Learning
by: Bhatnagar, Shubhang, et al.
Published: (2024)
by: Bhatnagar, Shubhang, et al.
Published: (2024)
Potential Field Based Deep Metric Learning
by: Bhatnagar, Shubhang, et al.
Published: (2024)
by: Bhatnagar, Shubhang, et al.
Published: (2024)
Rethinking Prompting Strategies for Multi-Label Recognition with Partial Annotations
by: Rawlekar, Samyak, et al.
Published: (2024)
by: Rawlekar, Samyak, et al.
Published: (2024)
Unsupervised Interpretable Basis Extraction for Concept-Based Visual Explanations
by: Doumanoglou, Alexandros, et al.
Published: (2023)
by: Doumanoglou, Alexandros, et al.
Published: (2023)
Technical Note: Defining and Quantifying AND-OR Interactions for Faithful and Concise Explanation of DNNs
by: Li, Mingjie, et al.
Published: (2023)
by: Li, Mingjie, et al.
Published: (2023)
FaCT: Faithful Concept Traces for Explaining Neural Network Decisions
by: Parchami-Araghi, Amin, et al.
Published: (2025)
by: Parchami-Araghi, Amin, et al.
Published: (2025)
DEAL: Disentangle and Localize Concept-level Explanations for VLMs
by: Li, Tang, et al.
Published: (2024)
by: Li, Tang, et al.
Published: (2024)
Statistically Significant Concept-based Explanation of Image Classifiers via Model Knockoffs
by: Xu, Kaiwen, et al.
Published: (2023)
by: Xu, Kaiwen, et al.
Published: (2023)
LUQ: Layerwise Ultra-Low Bit Quantization for Multimodal Large Language Models
by: Bhatnagar, Shubhang, et al.
Published: (2025)
by: Bhatnagar, Shubhang, et al.
Published: (2025)
Improving Diffusion-Based Image Editing Faithfulness via Guidance and Scheduling
by: Cho, Hansam, et al.
Published: (2025)
by: Cho, Hansam, et al.
Published: (2025)
FaithFill: Faithful Inpainting for Object Completion Using a Single Reference Image
by: Mallick, Rupayan, et al.
Published: (2024)
by: Mallick, Rupayan, et al.
Published: (2024)
Finding Distributed Object-Centric Properties in Self-Supervised Transformers
by: Rawlekar, Samyak, et al.
Published: (2026)
by: Rawlekar, Samyak, et al.
Published: (2026)
Improving Multi-label Recognition using Class Co-Occurrence Probabilities
by: Rawlekar, Samyak, et al.
Published: (2024)
by: Rawlekar, Samyak, et al.
Published: (2024)
Improving Interpretation Faithfulness for Vision Transformers
by: Hu, Lijie, et al.
Published: (2023)
by: Hu, Lijie, et al.
Published: (2023)
Concept-Based Unsupervised Domain Adaptation
by: Xu, Xinyue, et al.
Published: (2025)
by: Xu, Xinyue, et al.
Published: (2025)
Energy-Based Concept Bottleneck Models: Unifying Prediction, Concept Intervention, and Probabilistic Interpretations
by: Xu, Xinyue, et al.
Published: (2024)
by: Xu, Xinyue, et al.
Published: (2024)
When Bits Break Recourse: Counterfactual-Faithful Quantization
by: Yahyati, Chaymae, et al.
Published: (2026)
by: Yahyati, Chaymae, et al.
Published: (2026)
Explanation Bottleneck Models
by: Yamaguchi, Shin'ya, et al.
Published: (2024)
by: Yamaguchi, Shin'ya, et al.
Published: (2024)
Concept-Based Mechanistic Interpretability Using Structured Knowledge Graphs
by: Chorna, Sofiia, et al.
Published: (2025)
by: Chorna, Sofiia, et al.
Published: (2025)
Restyling Unsupervised Concept Based Interpretable Networks with Generative Models
by: Parekh, Jayneel, et al.
Published: (2024)
by: Parekh, Jayneel, et al.
Published: (2024)
Beyond Scalars: Concept-Based Alignment Analysis in Vision Transformers
by: Vielhaben, Johanna, et al.
Published: (2024)
by: Vielhaben, Johanna, et al.
Published: (2024)
Guaranteed Optimal Compositional Explanations for Neurons
by: La Rosa, Biagio, et al.
Published: (2025)
by: La Rosa, Biagio, et al.
Published: (2025)
MEGL: Multimodal Explanation-Guided Learning
by: Zhang, Yifei, et al.
Published: (2024)
by: Zhang, Yifei, et al.
Published: (2024)
A Geometric Unification of Concept Learning with Concept Cones
by: Rocchi--Henry, Alexandre, et al.
Published: (2025)
by: Rocchi--Henry, Alexandre, et al.
Published: (2025)
Concept Arithmetics for Circumventing Concept Inhibition in Diffusion Models
by: Petsiuk, Vitali, et al.
Published: (2024)
by: Petsiuk, Vitali, et al.
Published: (2024)
Learning Concept-Based Causal Transition and Symbolic Reasoning for Visual Planning
by: Qian, Yilue, et al.
Published: (2023)
by: Qian, Yilue, et al.
Published: (2023)
On Spectral Properties of Gradient-based Explanation Methods
by: Mehrpanah, Amir, et al.
Published: (2025)
by: Mehrpanah, Amir, et al.
Published: (2025)
Open Vocabulary Compositional Explanations for Neuron Alignment
by: La Rosa, Biagio, et al.
Published: (2025)
by: La Rosa, Biagio, et al.
Published: (2025)
Post-Hoc Concept Disentanglement: From Correlated to Isolated Concept Representations
by: Erogullari, Eren, et al.
Published: (2025)
by: Erogullari, Eren, et al.
Published: (2025)
Improving Intervention Efficacy via Concept Realignment in Concept Bottleneck Models
by: Singhi, Nishad, et al.
Published: (2024)
by: Singhi, Nishad, et al.
Published: (2024)
Concept Weaver: Enabling Multi-Concept Fusion in Text-to-Image Models
by: Kwon, Gihyun, et al.
Published: (2024)
by: Kwon, Gihyun, et al.
Published: (2024)
ICED: Concept-level Machine Unlearning via Interpretable Concept Decomposition
by: Lin, Shen, et al.
Published: (2026)
by: Lin, Shen, et al.
Published: (2026)
Low-cost Robust Night-time Aerial Material Segmentation through Hyperspectral Data and Sparse Spatio-Temporal Learning
by: Bajaj, Chandrajit, et al.
Published: (2024)
by: Bajaj, Chandrajit, et al.
Published: (2024)
Discover-then-Name: Task-Agnostic Concept Bottlenecks via Automated Concept Discovery
by: Rao, Sukrut, et al.
Published: (2024)
by: Rao, Sukrut, et al.
Published: (2024)
ConceptPrune: Concept Editing in Diffusion Models via Skilled Neuron Pruning
by: Chavhan, Ruchika, et al.
Published: (2024)
by: Chavhan, Ruchika, et al.
Published: (2024)
Pixel-level Certified Explanations via Randomized Smoothing
by: Anani, Alaa, et al.
Published: (2025)
by: Anani, Alaa, et al.
Published: (2025)
Towards Multi-dimensional Explanation Alignment for Medical Classification
by: Hu, Lijie, et al.
Published: (2024)
by: Hu, Lijie, et al.
Published: (2024)
Studying How to Efficiently and Effectively Guide Models with Explanations
by: Rao, Sukrut, et al.
Published: (2023)
by: Rao, Sukrut, et al.
Published: (2023)
Relevant Irrelevance: Generating Alterfactual Explanations for Image Classifiers
by: Mertes, Silvan, et al.
Published: (2024)
by: Mertes, Silvan, et al.
Published: (2024)
Similar Items
-
On the Complexity-Faithfulness Trade-off of Gradient-Based Explanations
by: Mehrpanah, Amir, et al.
Published: (2025) -
Piecewise-Linear Manifolds for Deep Metric Learning
by: Bhatnagar, Shubhang, et al.
Published: (2024) -
Potential Field Based Deep Metric Learning
by: Bhatnagar, Shubhang, et al.
Published: (2024) -
Rethinking Prompting Strategies for Multi-Label Recognition with Partial Annotations
by: Rawlekar, Samyak, et al.
Published: (2024) -
Unsupervised Interpretable Basis Extraction for Concept-Based Visual Explanations
by: Doumanoglou, Alexandros, et al.
Published: (2023)