A Dual-Perspective Approach to Evaluating Feature Attribution Methods
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Yawei, Zhang, Yang, Kawaguchi, Kenji, Khakzar, Ashkan, Bischl, Bernd, Rezaei, Mina |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AttributionLab: Faithfulness of Feature Attribution Under Controllable Environments
by: Zhang, Yang, et al.
Published: (2023)
by: Zhang, Yang, et al.
Published: (2023)
Calibrating LLMs with Information-Theoretic Evidential Deep Learning
by: Li, Yawei, et al.
Published: (2025)
by: Li, Yawei, et al.
Published: (2025)
FinerCut: Finer-grained Interpretable Layer Pruning for Large Language Models
by: Zhang, Yang, et al.
Published: (2024)
by: Zhang, Yang, et al.
Published: (2024)
What Is Fairness? On the Role of Protected Attributes and Fictitious Worlds
by: Bothmann, Ludwig, et al.
Published: (2022)
by: Bothmann, Ludwig, et al.
Published: (2022)
Analyzing Error Sources in Global Feature Effect Estimation
by: Heiß, Timo, et al.
Published: (2026)
by: Heiß, Timo, et al.
Published: (2026)
Quantifying Feature Space Universality Across Large Language Models via Sparse Autoencoders
by: Lan, Michael, et al.
Published: (2024)
by: Lan, Michael, et al.
Published: (2024)
On Discprecncies between Perturbation Evaluations of Graph Neural Network Attributions
by: Rezaei, Razieh, et al.
Published: (2024)
by: Rezaei, Razieh, et al.
Published: (2024)
Rethinking Robustness: A New Approach to Evaluating Feature Attribution Methods
by: Kiourti, Panagiota, et al.
Published: (2025)
by: Kiourti, Panagiota, et al.
Published: (2025)
Structured Credal Learning
by: Venkatesh, Varun, et al.
Published: (2026)
by: Venkatesh, Varun, et al.
Published: (2026)
A Comprehensive Machine Learning Framework for Heart Disease Prediction: Performance Evaluation and Future Perspectives
by: Lamir, Ali Azimi, et al.
Published: (2025)
by: Lamir, Ali Azimi, et al.
Published: (2025)
The Cognitive Revolution in Interpretability: From Explaining Behavior to Interpreting Representations and Algorithms
by: Davies, Adam, et al.
Published: (2024)
by: Davies, Adam, et al.
Published: (2024)
Explanation Space: A New Perspective into Time Series Interpretability
by: Rezaei, Shahbaz, et al.
Published: (2024)
by: Rezaei, Shahbaz, et al.
Published: (2024)
One Wave To Explain Them All: A Unifying Perspective On Feature Attribution
by: Kasmi, Gabriel, et al.
Published: (2024)
by: Kasmi, Gabriel, et al.
Published: (2024)
Referee Can Play: An Alternative Approach to Conditional Generation via Model Inversion
by: Liu, Xuantong, et al.
Published: (2024)
by: Liu, Xuantong, et al.
Published: (2024)
Latent Guard: a Safety Framework for Text-to-image Generation
by: Liu, Runtao, et al.
Published: (2024)
by: Liu, Runtao, et al.
Published: (2024)
LoReUn: Data Itself Implicitly Provides Cues to Improve Machine Unlearning
by: Li, Xiang, et al.
Published: (2025)
by: Li, Xiang, et al.
Published: (2025)
On Training Survival Models with Scoring Rules
by: Kopper, Philipp, et al.
Published: (2024)
by: Kopper, Philipp, et al.
Published: (2024)
RelP: Faithful and Efficient Circuit Discovery in Language Models via Relevance Patching
by: Jafari, Farnoush Rezaei, et al.
Published: (2025)
by: Jafari, Farnoush Rezaei, et al.
Published: (2025)
TABFAIRGDT: A Fast Fair Tabular Data Generator using Autoregressive Decision Trees
by: Panagiotou, Emmanouil, et al.
Published: (2025)
by: Panagiotou, Emmanouil, et al.
Published: (2025)
Impossibility Theorems for Feature Attribution
by: Bilodeau, Blair, et al.
Published: (2022)
by: Bilodeau, Blair, et al.
Published: (2022)
Uncertainty-aware Pseudo-label Selection for Positive-Unlabeled Learning
by: Dorigatti, Emilio, et al.
Published: (2022)
by: Dorigatti, Emilio, et al.
Published: (2022)
Attribution Explanations for Deep Neural Networks: A Theoretical Perspective
by: Deng, Huiqi, et al.
Published: (2025)
by: Deng, Huiqi, et al.
Published: (2025)
FreqX: Analyze the Attribution Methods in Another Domain
by: Liu, Zechen, et al.
Published: (2024)
by: Liu, Zechen, et al.
Published: (2024)
Why Do Class-Dependent Evaluation Effects Occur with Time Series Feature Attributions? A Synthetic Data Investigation
by: Baer, Gregor, et al.
Published: (2025)
by: Baer, Gregor, et al.
Published: (2025)
Correlation-Aware Feature Attribution Based Explainable AI
by: Sengupta, Poushali, et al.
Published: (2025)
by: Sengupta, Poushali, et al.
Published: (2025)
Probabilistic Stability Guarantees for Feature Attributions
by: Jin, Helen, et al.
Published: (2025)
by: Jin, Helen, et al.
Published: (2025)
Distribution-Based Feature Attribution for Explaining the Predictions of Any Classifier
by: Li, Xinpeng, et al.
Published: (2025)
by: Li, Xinpeng, et al.
Published: (2025)
xaitimesynth: A Python Package for Evaluating Attribution Methods for Time Series with Synthetic Ground Truth
by: Baer, Gregor
Published: (2026)
by: Baer, Gregor
Published: (2026)
On the Properties of Feature Attribution for Supervised Contrastive Learning
by: Arrighi, Leonardo, et al.
Published: (2026)
by: Arrighi, Leonardo, et al.
Published: (2026)
Simple Hierarchical Planning with Diffusion
by: Chen, Chang, et al.
Published: (2024)
by: Chen, Chang, et al.
Published: (2024)
A Polynomial-Time Axiomatic Alternative to SHAP for Feature Attribution
by: Hiraki, Kazuhiro, et al.
Published: (2026)
by: Hiraki, Kazuhiro, et al.
Published: (2026)
Memory-Efficient Gradient Unrolling for Large-Scale Bi-level Optimization
by: Shen, Qianli, et al.
Published: (2024)
by: Shen, Qianli, et al.
Published: (2024)
ChaosMining: A Benchmark to Evaluate Post-Hoc Local Attribution Methods in Low SNR Environments
by: Shi, Ge, et al.
Published: (2024)
by: Shi, Ge, et al.
Published: (2024)
ReAGent: A Model-agnostic Feature Attribution Method for Generative Language Models
by: Zhao, Zhixue, et al.
Published: (2024)
by: Zhao, Zhixue, et al.
Published: (2024)
Sparse, Efficient and Explainable Data Attribution with DualXDA
by: Yolcu, Galip Ümit, et al.
Published: (2024)
by: Yolcu, Galip Ümit, et al.
Published: (2024)
Set-based Meta-Interpolation for Few-Task Meta-Learning
by: Lee, Seanie, et al.
Published: (2022)
by: Lee, Seanie, et al.
Published: (2022)
AttributionBench: How Hard is Automatic Attribution Evaluation?
by: Li, Yifei, et al.
Published: (2024)
by: Li, Yifei, et al.
Published: (2024)
AlignGuard: Scalable Safety Alignment for Text-to-Image Generation
by: Liu, Runtao, et al.
Published: (2024)
by: Liu, Runtao, et al.
Published: (2024)
COVID-19 Probability Prediction Using Machine Learning: An Infectious Approach
by: Ilani, Mohsen Asghari, et al.
Published: (2024)
by: Ilani, Mohsen Asghari, et al.
Published: (2024)
RDKV: Rate-Distortion Bit Allocation for Joint Eviction and Quantization of the KV Cache
by: Zhang, Junkai, et al.
Published: (2026)
by: Zhang, Junkai, et al.
Published: (2026)
Similar Items
-
AttributionLab: Faithfulness of Feature Attribution Under Controllable Environments
by: Zhang, Yang, et al.
Published: (2023) -
Calibrating LLMs with Information-Theoretic Evidential Deep Learning
by: Li, Yawei, et al.
Published: (2025) -
FinerCut: Finer-grained Interpretable Layer Pruning for Large Language Models
by: Zhang, Yang, et al.
Published: (2024) -
What Is Fairness? On the Role of Protected Attributes and Fictitious Worlds
by: Bothmann, Ludwig, et al.
Published: (2022) -
Analyzing Error Sources in Global Feature Effect Estimation
by: Heiß, Timo, et al.
Published: (2026)