Statistically Significant Concept-based Explanation of Image Classifiers via Model Knockoffs
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Kaiwen, Fukuchi, Kazuto, Akimoto, Youhei, Sakuma, Jun |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CAMRI Loss: Improving Recall of a Specific Class without Sacrificing Accuracy
by: Nishiyama, Daiki, et al.
Published: (2022)
by: Nishiyama, Daiki, et al.
Published: (2022)
Relevant Irrelevance: Generating Alterfactual Explanations for Image Classifiers
by: Mertes, Silvan, et al.
Published: (2024)
by: Mertes, Silvan, et al.
Published: (2024)
Robust Deep Reinforcement Learning against Adversarial Behavior Manipulation
by: Yamabe, Shojiro, et al.
Published: (2024)
by: Yamabe, Shojiro, et al.
Published: (2024)
Accurate Explanation Model for Image Classifiers using Class Association Embedding
by: Xie, Ruitao, et al.
Published: (2024)
by: Xie, Ruitao, et al.
Published: (2024)
Measuring the (Un)Faithfulness of Concept-Based Explanations
by: Kumar, Shubham, et al.
Published: (2025)
by: Kumar, Shubham, et al.
Published: (2025)
Multiple Different Black Box Explanations for Image Classifiers
by: Chockler, Hana, et al.
Published: (2023)
by: Chockler, Hana, et al.
Published: (2023)
DEAL: Disentangle and Localize Concept-level Explanations for VLMs
by: Li, Tang, et al.
Published: (2024)
by: Li, Tang, et al.
Published: (2024)
Unsupervised Interpretable Basis Extraction for Concept-Based Visual Explanations
by: Doumanoglou, Alexandros, et al.
Published: (2023)
by: Doumanoglou, Alexandros, et al.
Published: (2023)
Concept Unlearning via Cross-Attention Activation Projection for Diffusion Models
by: Moon, Saemi, et al.
Published: (2026)
by: Moon, Saemi, et al.
Published: (2026)
TbExplain: A Text-based Explanation Method for Scene Classification Models with the Statistical Prediction Correction
by: Aminimehr, Amirhossein, et al.
Published: (2023)
by: Aminimehr, Amirhossein, et al.
Published: (2023)
Concept Weaver: Enabling Multi-Concept Fusion in Text-to-Image Models
by: Kwon, Gihyun, et al.
Published: (2024)
by: Kwon, Gihyun, et al.
Published: (2024)
HOLMES: HOLonym-MEronym based Semantic inspection for Convolutional Image Classifiers
by: Dibitonto, Francesco, et al.
Published: (2024)
by: Dibitonto, Francesco, et al.
Published: (2024)
Zero-Residual Concept Erasure via Progressive Alignment in Text-to-Image Model
by: Chen, Hongxu, et al.
Published: (2025)
by: Chen, Hongxu, et al.
Published: (2025)
Convergence rate of the (1+1)-evolution strategy on locally strongly convex functions with lipschitz continuous gradient
by: Morinaga, Daiki, et al.
Published: (2022)
by: Morinaga, Daiki, et al.
Published: (2022)
Instance-Level Trojan Attacks on Visual Question Answering via Adversarial Learning in Neuron Activation Space
by: Sun, Yuwei, et al.
Published: (2023)
by: Sun, Yuwei, et al.
Published: (2023)
Prompt Optimizer of Text-to-Image Diffusion Models for Abstract Concept Understanding
by: Fan, Zezhong, et al.
Published: (2024)
by: Fan, Zezhong, et al.
Published: (2024)
Explainability of Point Cloud Neural Networks Using SMILE: Statistical Model-Agnostic Interpretability with Local Explanations
by: Ahmadi, Seyed Mohammad, et al.
Published: (2024)
by: Ahmadi, Seyed Mohammad, et al.
Published: (2024)
Improving Intervention Efficacy via Concept Realignment in Concept Bottleneck Models
by: Singhi, Nishad, et al.
Published: (2024)
by: Singhi, Nishad, et al.
Published: (2024)
ICED: Concept-level Machine Unlearning via Interpretable Concept Decomposition
by: Lin, Shen, et al.
Published: (2026)
by: Lin, Shen, et al.
Published: (2026)
Interactive Medical Image Analysis with Concept-based Similarity Reasoning
by: Huy, Ta Duc, et al.
Published: (2025)
by: Huy, Ta Duc, et al.
Published: (2025)
DiG-IN: Diffusion Guidance for Investigating Networks -- Uncovering Classifier Differences Neuron Visualisations and Visual Counterfactual Explanations
by: Augustin, Maximilian, et al.
Published: (2023)
by: Augustin, Maximilian, et al.
Published: (2023)
LINE: LLM-based Iterative Neuron Explanations for Vision Models
by: Zaigrajew, Vladimir, et al.
Published: (2026)
by: Zaigrajew, Vladimir, et al.
Published: (2026)
Editing Massive Concepts in Text-to-Image Diffusion Models
by: Xiong, Tianwei, et al.
Published: (2024)
by: Xiong, Tianwei, et al.
Published: (2024)
ConceptPrune: Concept Editing in Diffusion Models via Skilled Neuron Pruning
by: Chavhan, Ruchika, et al.
Published: (2024)
by: Chavhan, Ruchika, et al.
Published: (2024)
Visual Explanations of Image-Text Representations via Multi-Modal Information Bottleneck Attribution
by: Wang, Ying, et al.
Published: (2023)
by: Wang, Ying, et al.
Published: (2023)
Understanding the (Extra-)Ordinary: Validating Deep Model Decisions with Prototypical Concept-based Explanations
by: Dreyer, Maximilian, et al.
Published: (2023)
by: Dreyer, Maximilian, et al.
Published: (2023)
Explanation Bottleneck Models
by: Yamaguchi, Shin'ya, et al.
Published: (2024)
by: Yamaguchi, Shin'ya, et al.
Published: (2024)
Leveraging Model Soups to Classify Intangible Cultural Heritage Images from the Mekong Delta
by: Tran, Quoc-Khang, et al.
Published: (2026)
by: Tran, Quoc-Khang, et al.
Published: (2026)
Towards Facilitated Fairness Assessment of AI-based Skin Lesion Classifiers Through GenAI-based Image Synthesis
by: Watanabe, Ko, et al.
Published: (2025)
by: Watanabe, Ko, et al.
Published: (2025)
Energy-Based Concept Bottleneck Models: Unifying Prediction, Concept Intervention, and Probabilistic Interpretations
by: Xu, Xinyue, et al.
Published: (2024)
by: Xu, Xinyue, et al.
Published: (2024)
On Spectral Properties of Gradient-based Explanation Methods
by: Mehrpanah, Amir, et al.
Published: (2025)
by: Mehrpanah, Amir, et al.
Published: (2025)
Sculpting Memory: Multi-Concept Forgetting in Diffusion Models via Dynamic Mask and Concept-Aware Optimization
by: Li, Gen, et al.
Published: (2025)
by: Li, Gen, et al.
Published: (2025)
ViGText: Deepfake Image Detection with Vision-Language Model Explanations and Graph Neural Networks
by: ALBarqawi, Ahmad, et al.
Published: (2025)
by: ALBarqawi, Ahmad, et al.
Published: (2025)
Identifying and Solving Conditional Image Leakage in Image-to-Video Diffusion Model
by: Zhao, Min, et al.
Published: (2024)
by: Zhao, Min, et al.
Published: (2024)
Sparse Autoencoder as a Zero-Shot Classifier for Concept Erasing in Text-to-Image Diffusion Models
by: Tian, Zhihua, et al.
Published: (2025)
by: Tian, Zhihua, et al.
Published: (2025)
Investigating the Corruption Robustness of Image Classifiers with Random Lp-norm Corruptions
by: Siedel, Georg, et al.
Published: (2023)
by: Siedel, Georg, et al.
Published: (2023)
Controllable Unlearning for Image-to-Image Generative Models via $\varepsilon$-Constrained Optimization
by: Feng, Xiaohua, et al.
Published: (2024)
by: Feng, Xiaohua, et al.
Published: (2024)
Concept Arithmetics for Circumventing Concept Inhibition in Diffusion Models
by: Petsiuk, Vitali, et al.
Published: (2024)
by: Petsiuk, Vitali, et al.
Published: (2024)
Deformable ProtoPNet: An Interpretable Image Classifier Using Deformable Prototypes
by: Donnelly, Jon, et al.
Published: (2021)
by: Donnelly, Jon, et al.
Published: (2021)
Visual-TCAV: Concept-based Attribution and Saliency Maps for Post-hoc Explainability in Image Classification
by: De Santis, Antonio, et al.
Published: (2024)
by: De Santis, Antonio, et al.
Published: (2024)
Similar Items
-
CAMRI Loss: Improving Recall of a Specific Class without Sacrificing Accuracy
by: Nishiyama, Daiki, et al.
Published: (2022) -
Relevant Irrelevance: Generating Alterfactual Explanations for Image Classifiers
by: Mertes, Silvan, et al.
Published: (2024) -
Robust Deep Reinforcement Learning against Adversarial Behavior Manipulation
by: Yamabe, Shojiro, et al.
Published: (2024) -
Accurate Explanation Model for Image Classifiers using Class Association Embedding
by: Xie, Ruitao, et al.
Published: (2024) -
Measuring the (Un)Faithfulness of Concept-Based Explanations
by: Kumar, Shubham, et al.
Published: (2025)