Explanation-Driven Counterfactual Testing for Faithfulness in Vision-Language Model Explanations
Fuente:
arXiv
Saved in:
| Main Authors: | Ding, Sihao, Vasa, Santosh, Ramadwar, Aditi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AutoVDC: Automated Vision Data Cleaning Using Vision-Language Models
by: Vasa, Santosh, et al.
Published: (2025)
by: Vasa, Santosh, et al.
Published: (2025)
Rethinking Visual Counterfactual Explanations Through Region Constraint
by: Sobieski, Bartlomiej, et al.
Published: (2024)
by: Sobieski, Bartlomiej, et al.
Published: (2024)
SAM-Sode: Towards Faithful Explanations for Tiny Bacteria Detection
by: Tan, Wanying, et al.
Published: (2026)
by: Tan, Wanying, et al.
Published: (2026)
V-CECE: Visual Counterfactual Explanations via Conceptual Edits
by: Spanos, Nikolaos, et al.
Published: (2025)
by: Spanos, Nikolaos, et al.
Published: (2025)
Faithful Counterfactual Visual Explanations (FCVE)
by: Khan, Bismillah, et al.
Published: (2025)
by: Khan, Bismillah, et al.
Published: (2025)
Measuring the (Un)Faithfulness of Concept-Based Explanations
by: Kumar, Shubham, et al.
Published: (2025)
by: Kumar, Shubham, et al.
Published: (2025)
Local-to-Global Logical Explanations for Deep Vision Models
by: Vasu, Bhavan, et al.
Published: (2026)
by: Vasu, Bhavan, et al.
Published: (2026)
On the Complexity-Faithfulness Trade-off of Gradient-Based Explanations
by: Mehrpanah, Amir, et al.
Published: (2025)
by: Mehrpanah, Amir, et al.
Published: (2025)
U-CECE: A Universal Multi-Resolution Framework for Conceptual Counterfactual Explanations
by: Dimitriou, Angeliki, et al.
Published: (2026)
by: Dimitriou, Angeliki, et al.
Published: (2026)
Why Are You Wrong? Counterfactual Explanations for Language Grounding with 3D Objects
by: Preintner, Tobias, et al.
Published: (2025)
by: Preintner, Tobias, et al.
Published: (2025)
On the Faithfulness of Vision Transformer Explanations
by: Wu, Junyi, et al.
Published: (2024)
by: Wu, Junyi, et al.
Published: (2024)
Technical Note: Defining and Quantifying AND-OR Interactions for Faithful and Concise Explanation of DNNs
by: Li, Mingjie, et al.
Published: (2023)
by: Li, Mingjie, et al.
Published: (2023)
Musketeer: Joint Training for Multi-task Vision Language Model with Task Explanation Prompts
by: Zhang, Zhaoyang, et al.
Published: (2023)
by: Zhang, Zhaoyang, et al.
Published: (2023)
Scaling Vision Models Does Not Consistently Improve Localisation-Based Explanation Quality
by: Cedro, Mateusz, et al.
Published: (2026)
by: Cedro, Mateusz, et al.
Published: (2026)
From Pixels to Explanations: Interpretable Diabetic Retinopathy Grading with CNN-Transformer Ensembles, Visual Explainability and Vision-Language Models
by: Khokhar, Pir Bakhsh, et al.
Published: (2026)
by: Khokhar, Pir Bakhsh, et al.
Published: (2026)
Uncovering Bias in Large Vision-Language Models with Counterfactuals
by: Howard, Phillip, et al.
Published: (2024)
by: Howard, Phillip, et al.
Published: (2024)
LINE: LLM-based Iterative Neuron Explanations for Vision Models
by: Zaigrajew, Vladimir, et al.
Published: (2026)
by: Zaigrajew, Vladimir, et al.
Published: (2026)
SocialCounterfactuals: Probing and Mitigating Intersectional Social Biases in Vision-Language Models with Counterfactual Examples
by: Howard, Phillip, et al.
Published: (2023)
by: Howard, Phillip, et al.
Published: (2023)
Explanation Bottleneck Models
by: Yamaguchi, Shin'ya, et al.
Published: (2024)
by: Yamaguchi, Shin'ya, et al.
Published: (2024)
Spatially Grounded Explanations in Vision Language Models for Document Visual Question Answering
by: Lagos, Maximiliano Hormazábal, et al.
Published: (2025)
by: Lagos, Maximiliano Hormazábal, et al.
Published: (2025)
Attention Guided CAM: Visual Explanations of Vision Transformer Guided by Self-Attention
by: Leem, Saebom, et al.
Published: (2024)
by: Leem, Saebom, et al.
Published: (2024)
ViGText: Deepfake Image Detection with Vision-Language Model Explanations and Graph Neural Networks
by: ALBarqawi, Ahmad, et al.
Published: (2025)
by: ALBarqawi, Ahmad, et al.
Published: (2025)
Probabilistic Conceptual Explainers: Trustworthy Conceptual Explanations for Vision Foundation Models
by: Wang, Hengyi, et al.
Published: (2024)
by: Wang, Hengyi, et al.
Published: (2024)
Activation Steering Meets Preference Optimization: Defense Against Jailbreaks in Vision Language Models
by: Wu, Sihao, et al.
Published: (2025)
by: Wu, Sihao, et al.
Published: (2025)
Causal-Adapter: Taming Text-to-Image Diffusion for Faithful Counterfactual Generation
by: Tong, Lei, et al.
Published: (2025)
by: Tong, Lei, et al.
Published: (2025)
XBench: A Comprehensive Benchmark for Visual-Language Explanations in Chest Radiography
by: Luo, Haozhe, et al.
Published: (2025)
by: Luo, Haozhe, et al.
Published: (2025)
LangXAI: Integrating Large Vision Models for Generating Textual Explanations to Enhance Explainability in Visual Perception Tasks
by: Nguyen, Truong Thanh Hung, et al.
Published: (2024)
by: Nguyen, Truong Thanh Hung, et al.
Published: (2024)
On The Coherence of Quantitative Evaluation of Visual Explanations
by: Vandersmissen, Benjamin, et al.
Published: (2023)
by: Vandersmissen, Benjamin, et al.
Published: (2023)
Diffexplainer: Towards Cross-modal Global Explanations with Diffusion Models
by: Pennisi, Matteo, et al.
Published: (2024)
by: Pennisi, Matteo, et al.
Published: (2024)
FaithSCAN: Model-Driven Single-Pass Hallucination Detection for Faithful Visual Question Answering
by: Tong, Chaodong, et al.
Published: (2026)
by: Tong, Chaodong, et al.
Published: (2026)
Winsor-CAM: Human-Tunable Visual Explanations from Deep Networks via Layer-Wise Winsorization
by: Wall, Casey, et al.
Published: (2025)
by: Wall, Casey, et al.
Published: (2025)
What Helps---and What Hurts: Bidirectional Explanations for Vision Transformers
by: Su, Qin, et al.
Published: (2026)
by: Su, Qin, et al.
Published: (2026)
Noise-Free Explanation for Driving Action Prediction
by: Zhu, Hongbo, et al.
Published: (2024)
by: Zhu, Hongbo, et al.
Published: (2024)
MLLM-based Textual Explanations for Face Comparison
by: Sony, Redwan, et al.
Published: (2026)
by: Sony, Redwan, et al.
Published: (2026)
Accurate Explanation Model for Image Classifiers using Class Association Embedding
by: Xie, Ruitao, et al.
Published: (2024)
by: Xie, Ruitao, et al.
Published: (2024)
Case-Enhanced Vision Transformer: Improving Explanations of Image Similarity with a ViT-based Similarity Metric
by: Zhao, Ziwei, et al.
Published: (2024)
by: Zhao, Ziwei, et al.
Published: (2024)
DiG-IN: Diffusion Guidance for Investigating Networks -- Uncovering Classifier Differences Neuron Visualisations and Visual Counterfactual Explanations
by: Augustin, Maximilian, et al.
Published: (2023)
by: Augustin, Maximilian, et al.
Published: (2023)
Sufficient, Necessary and Complete Causal Explanations in Image Classification
by: Kelly, David A, et al.
Published: (2025)
by: Kelly, David A, et al.
Published: (2025)
Generating Part-Based Global Explanations Via Correspondence
by: Rathore, Kunal, et al.
Published: (2025)
by: Rathore, Kunal, et al.
Published: (2025)
From Attribution to Action: Jointly ALIGNing Predictions and Explanations
by: Hong, Dongsheng, et al.
Published: (2025)
by: Hong, Dongsheng, et al.
Published: (2025)
Similar Items
-
AutoVDC: Automated Vision Data Cleaning Using Vision-Language Models
by: Vasa, Santosh, et al.
Published: (2025) -
Rethinking Visual Counterfactual Explanations Through Region Constraint
by: Sobieski, Bartlomiej, et al.
Published: (2024) -
SAM-Sode: Towards Faithful Explanations for Tiny Bacteria Detection
by: Tan, Wanying, et al.
Published: (2026) -
V-CECE: Visual Counterfactual Explanations via Conceptual Edits
by: Spanos, Nikolaos, et al.
Published: (2025) -
Faithful Counterfactual Visual Explanations (FCVE)
by: Khan, Bismillah, et al.
Published: (2025)