One Wave To Explain Them All: A Unifying Perspective On Feature Attribution
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kasmi, Gabriel, Brunetto, Amandine, Fel, Thomas, Parekh, Jayneel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Distribution-Based Feature Attribution for Explaining the Predictions of Any Classifier
von: Li, Xinpeng, et al.
Veröffentlicht: (2025)
von: Li, Xinpeng, et al.
Veröffentlicht: (2025)
GNN Explanations that do not Explain and How to find Them
von: Azzolin, Steve, et al.
Veröffentlicht: (2026)
von: Azzolin, Steve, et al.
Veröffentlicht: (2026)
A Concept-Based Explainability Framework for Large Multimodal Models
von: Parekh, Jayneel, et al.
Veröffentlicht: (2024)
von: Parekh, Jayneel, et al.
Veröffentlicht: (2024)
Restyling Unsupervised Concept Based Interpretable Networks with Generative Models
von: Parekh, Jayneel, et al.
Veröffentlicht: (2024)
von: Parekh, Jayneel, et al.
Veröffentlicht: (2024)
A Dual-Perspective Approach to Evaluating Feature Attribution Methods
von: Li, Yawei, et al.
Veröffentlicht: (2023)
von: Li, Yawei, et al.
Veröffentlicht: (2023)
Unifying Perplexing Behaviors in Modified BP Attributions through Alignment Perspective
von: Zheng, Guanhua, et al.
Veröffentlicht: (2025)
von: Zheng, Guanhua, et al.
Veröffentlicht: (2025)
Learning to Steer: Input-dependent Steering for Multimodal LLMs
von: Parekh, Jayneel, et al.
Veröffentlicht: (2025)
von: Parekh, Jayneel, et al.
Veröffentlicht: (2025)
When Prompts Override Vision: Prompt-Induced Hallucinations in LVLMs
von: Khayatan, Pegah, et al.
Veröffentlicht: (2026)
von: Khayatan, Pegah, et al.
Veröffentlicht: (2026)
One Filter to Deploy Them All: Robust Safety for Quadrupedal Navigation in Unknown Environments
von: Lin, Albert, et al.
Veröffentlicht: (2024)
von: Lin, Albert, et al.
Veröffentlicht: (2024)
One Router to Route Them All: Homogeneous Expert Routing for Heterogeneous Graph Transformers
von: Shakirov, Georgiy, et al.
Veröffentlicht: (2025)
von: Shakirov, Georgiy, et al.
Veröffentlicht: (2025)
Explaining Time Series Classification Predictions via Causal Attributions
von: Alcaraz, Juan Miguel Lopez, et al.
Veröffentlicht: (2024)
von: Alcaraz, Juan Miguel Lopez, et al.
Veröffentlicht: (2024)
Drop the Act: Probe-Filtered RL for Faithful Chain-of-Thought Reasoning
von: Parekh, Swapnil
Veröffentlicht: (2026)
von: Parekh, Swapnil
Veröffentlicht: (2026)
Thinking Wrong in Silence: Backdoor Attacks on Continuous Latent Reasoning
von: Parekh, Swapnil
Veröffentlicht: (2026)
von: Parekh, Swapnil
Veröffentlicht: (2026)
One-Dimensional Adapter to Rule Them All: Concepts, Diffusion Models and Erasing Applications
von: Lyu, Mengyao, et al.
Veröffentlicht: (2023)
von: Lyu, Mengyao, et al.
Veröffentlicht: (2023)
Impossibility Theorems for Feature Attribution
von: Bilodeau, Blair, et al.
Veröffentlicht: (2022)
von: Bilodeau, Blair, et al.
Veröffentlicht: (2022)
On the Overlooked Pitfalls of Weight Decay and How to Mitigate Them: A Gradient-Norm Perspective
von: Xie, Zeke, et al.
Veröffentlicht: (2020)
von: Xie, Zeke, et al.
Veröffentlicht: (2020)
Attributions All the Way Down? The Metagame of Interpretability
von: Baniecki, Hubert, et al.
Veröffentlicht: (2026)
von: Baniecki, Hubert, et al.
Veröffentlicht: (2026)
Probabilistic Stability Guarantees for Feature Attributions
von: Jin, Helen, et al.
Veröffentlicht: (2025)
von: Jin, Helen, et al.
Veröffentlicht: (2025)
Sparks of Explainability: Recent Advancements in Explaining Large Vision Models
von: Fel, Thomas
Veröffentlicht: (2025)
von: Fel, Thomas
Veröffentlicht: (2025)
Projecting Assumptions: The Duality Between Sparse Autoencoders and Concept Geometry
von: Hindupur, Sai Sumedh R., et al.
Veröffentlicht: (2025)
von: Hindupur, Sai Sumedh R., et al.
Veröffentlicht: (2025)
Who's the (Multi-)Fairest of Them All: Rethinking Interpolation-Based Data Augmentation Through the Lens of Multicalibration
von: Halevy, Karina, et al.
Veröffentlicht: (2024)
von: Halevy, Karina, et al.
Veröffentlicht: (2024)
Helpful or Harmful Data? Fine-tuning-free Shapley Attribution for Explaining Language Model Predictions
von: Wang, Jingtan, et al.
Veröffentlicht: (2024)
von: Wang, Jingtan, et al.
Veröffentlicht: (2024)
Explaining Time Series Classifiers with PHAR: Rule Extraction and Fusion from Post-hoc Attributions
von: Mozolewski, Maciej, et al.
Veröffentlicht: (2025)
von: Mozolewski, Maciej, et al.
Veröffentlicht: (2025)
Explaining Categorical Feature Interactions Using Graph Covariance and LLMs
von: Shen, Cencheng, et al.
Veröffentlicht: (2025)
von: Shen, Cencheng, et al.
Veröffentlicht: (2025)
A Geometric Unification of Concept Learning with Concept Cones
von: Rocchi--Henry, Alexandre, et al.
Veröffentlicht: (2025)
von: Rocchi--Henry, Alexandre, et al.
Veröffentlicht: (2025)
On the Properties of Feature Attribution for Supervised Contrastive Learning
von: Arrighi, Leonardo, et al.
Veröffentlicht: (2026)
von: Arrighi, Leonardo, et al.
Veröffentlicht: (2026)
Unifying Attribution-Based Explanations Using Functional Decomposition
von: Gevaert, Arne, et al.
Veröffentlicht: (2024)
von: Gevaert, Arne, et al.
Veröffentlicht: (2024)
Feature-Function Curvature Analysis: A Geometric Framework for Explaining Differentiable Models
von: Najafi, Hamed, et al.
Veröffentlicht: (2025)
von: Najafi, Hamed, et al.
Veröffentlicht: (2025)
A Polynomial-Time Axiomatic Alternative to SHAP for Feature Attribution
von: Hiraki, Kazuhiro, et al.
Veröffentlicht: (2026)
von: Hiraki, Kazuhiro, et al.
Veröffentlicht: (2026)
ABE: A Unified Framework for Robust and Faithful Attribution-Based Explainability
von: Zhu, Zhiyu, et al.
Veröffentlicht: (2025)
von: Zhu, Zhiyu, et al.
Veröffentlicht: (2025)
Attribution Explanations for Deep Neural Networks: A Theoretical Perspective
von: Deng, Huiqi, et al.
Veröffentlicht: (2025)
von: Deng, Huiqi, et al.
Veröffentlicht: (2025)
CIRCUS: Circuit Consensus under Uncertainty via Stability Ensembles
von: Parekh, Swapnil
Veröffentlicht: (2026)
von: Parekh, Swapnil
Veröffentlicht: (2026)
Delta-XAI: A Unified Framework for Explaining Prediction Changes in Online Time Series Monitoring
von: Kim, Changhun, et al.
Veröffentlicht: (2025)
von: Kim, Changhun, et al.
Veröffentlicht: (2025)
On the explainable properties of 1-Lipschitz Neural Networks: An Optimal Transport Perspective
von: Serrurier, Mathieu, et al.
Veröffentlicht: (2022)
von: Serrurier, Mathieu, et al.
Veröffentlicht: (2022)
Correlation-Aware Feature Attribution Based Explainable AI
von: Sengupta, Poushali, et al.
Veröffentlicht: (2025)
von: Sengupta, Poushali, et al.
Veröffentlicht: (2025)
Learning Unified Distance Metric for Heterogeneous Attribute Data Clustering
von: Zhang, Yiqun, et al.
Veröffentlicht: (2026)
von: Zhang, Yiqun, et al.
Veröffentlicht: (2026)
Pruning By Explaining Revisited: Optimizing Attribution Methods to Prune CNNs and Transformers
von: Hatefi, Sayed Mohammad Vakilzadeh, et al.
Veröffentlicht: (2024)
von: Hatefi, Sayed Mohammad Vakilzadeh, et al.
Veröffentlicht: (2024)
All Random Features Representations are Equivalent
von: Sernau, Luke, et al.
Veröffentlicht: (2024)
von: Sernau, Luke, et al.
Veröffentlicht: (2024)
HMVI: Unifying Heterogeneous Attributes with Natural Neighbors for Missing Value Inference
von: Luo, Xiaopeng, et al.
Veröffentlicht: (2026)
von: Luo, Xiaopeng, et al.
Veröffentlicht: (2026)
One Filters All: A Generalist Filter for State Estimation
von: Liu, Shiqi, et al.
Veröffentlicht: (2025)
von: Liu, Shiqi, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Distribution-Based Feature Attribution for Explaining the Predictions of Any Classifier
von: Li, Xinpeng, et al.
Veröffentlicht: (2025) -
GNN Explanations that do not Explain and How to find Them
von: Azzolin, Steve, et al.
Veröffentlicht: (2026) -
A Concept-Based Explainability Framework for Large Multimodal Models
von: Parekh, Jayneel, et al.
Veröffentlicht: (2024) -
Restyling Unsupervised Concept Based Interpretable Networks with Generative Models
von: Parekh, Jayneel, et al.
Veröffentlicht: (2024) -
A Dual-Perspective Approach to Evaluating Feature Attribution Methods
von: Li, Yawei, et al.
Veröffentlicht: (2023)