Saved in:
| Main Authors: | Gjølbye, Anders, Haufe, Stefan, Hansen, Lars Kai |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2505.11210 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reasoning Models Don't Just Think Longer, They Move Differently
by: Gjølbye, Anders, et al.
Published: (2026)
by: Gjølbye, Anders, et al.
Published: (2026)
Missing-Data-Induced Phase Transitions in Spectral PLS for Multimodal Learning
by: Gjølbye, Anders, et al.
Published: (2026)
by: Gjølbye, Anders, et al.
Published: (2026)
Concept-based explainability for an EEG transformer model
by: Gjølbye, Anders, et al.
Published: (2023)
by: Gjølbye, Anders, et al.
Published: (2023)
Robustness of Visual Explanations to Common Data Augmentation
by: Tětková, Lenka, et al.
Published: (2023)
by: Tětková, Lenka, et al.
Published: (2023)
cc-Shapley: Measuring Multivariate Feature Importance Needs Causal Context
by: Martin, Jörg, et al.
Published: (2026)
by: Martin, Jörg, et al.
Published: (2026)
GECOBench: A Gender-Controlled Text Dataset and Benchmark for Quantifying Biases in Explanations
by: Wilming, Rick, et al.
Published: (2024)
by: Wilming, Rick, et al.
Published: (2024)
Benchmarking the Influence of Pre-training on Explanation Performance in MR Image Classification
by: Oliveira, Marta, et al.
Published: (2023)
by: Oliveira, Marta, et al.
Published: (2023)
The effect of whitening on explanation performance
by: Clark, Benedict, et al.
Published: (2026)
by: Clark, Benedict, et al.
Published: (2026)
Feature salience - not task-informativeness - drives machine learning model explanations
by: Clark, Benedict, et al.
Published: (2026)
by: Clark, Benedict, et al.
Published: (2026)
Generative clinical time series models trained on moderate amounts of patient data are privacy preserving
by: Zhumagambetov, Rustam, et al.
Published: (2026)
by: Zhumagambetov, Rustam, et al.
Published: (2026)
Large Vision Models Can Solve Mental Rotation Problems
by: Mason, Sebastian Ray, et al.
Published: (2025)
by: Mason, Sebastian Ray, et al.
Published: (2025)
SPEED: Scalable Preprocessing of EEG Data for Self-Supervised Learning
by: Gjølbye, Anders, et al.
Published: (2024)
by: Gjølbye, Anders, et al.
Published: (2024)
Backward Compatibility in Attributive Explanation and Enhanced Model Training Method
by: Matsuno, Ryuta
Published: (2024)
by: Matsuno, Ryuta
Published: (2024)
Beyond Attribution: Unified Concept-Level Explanations
by: Liu, Junhao, et al.
Published: (2024)
by: Liu, Junhao, et al.
Published: (2024)
Missingness Bias Calibration in Feature Attribution Explanations
by: Sridhar, Shailesh, et al.
Published: (2026)
by: Sridhar, Shailesh, et al.
Published: (2026)
Locally-Minimal Probabilistic Explanations
by: Izza, Yacine, et al.
Published: (2023)
by: Izza, Yacine, et al.
Published: (2023)
Model-Based Counterfactual Explanations Incorporating Feature Space Attributes for Tabular Data
by: Sumiya, Yuta, et al.
Published: (2024)
by: Sumiya, Yuta, et al.
Published: (2024)
Quotient Semivalues for False-Name-Resistant Data Attribution
by: Burnat, Florian A. D., et al.
Published: (2026)
by: Burnat, Florian A. D., et al.
Published: (2026)
Interpretability in Parameter Space: Minimizing Mechanistic Description Length with Attribution-based Parameter Decomposition
by: Braun, Dan, et al.
Published: (2025)
by: Braun, Dan, et al.
Published: (2025)
Enhancing Brain Source Reconstruction by Initializing 3D Neural Networks with Physical Inverse Solutions
by: Morik, Marco, et al.
Published: (2024)
by: Morik, Marco, et al.
Published: (2024)
Steering LLMs? Actually, Sparse Autoencoders can outperform simple baselines
by: Jørgensen, Mikkel Godsk, et al.
Published: (2026)
by: Jørgensen, Mikkel Godsk, et al.
Published: (2026)
Taming False Positives in Out-of-Distribution Detection with Human Feedback
by: Vishwakarma, Harit, et al.
Published: (2024)
by: Vishwakarma, Harit, et al.
Published: (2024)
Why Knowledge Distillation Works in Generative Models: A Minimal Working Explanation
by: Cha, Sungmin, et al.
Published: (2025)
by: Cha, Sungmin, et al.
Published: (2025)
Weak-to-Strong Generalization is Nearly Inevitable (in Linear Models)
by: Geng, Scott, et al.
Published: (2026)
by: Geng, Scott, et al.
Published: (2026)
Unifying Attribution-Based Explanations Using Functional Decomposition
by: Gevaert, Arne, et al.
Published: (2024)
by: Gevaert, Arne, et al.
Published: (2024)
Hypothesis Class Determines Explanation: Why Accurate Models Disagree on Feature Attribution
by: B, Thackshanaramana
Published: (2026)
by: B, Thackshanaramana
Published: (2026)
Counterfactual Explanations for Linear Optimization
by: Kurtz, Jannis, et al.
Published: (2024)
by: Kurtz, Jannis, et al.
Published: (2024)
When Explanations Lie: Why Many Modified BP Attributions Fail
by: Sixt, Leon, et al.
Published: (2019)
by: Sixt, Leon, et al.
Published: (2019)
Continual Adversarial Reinforcement Learning (CARL) of False Data Injection detection: forgetting and explainability
by: Aslami, Pooja, et al.
Published: (2024)
by: Aslami, Pooja, et al.
Published: (2024)
Counterfactual Explanations via Riemannian Latent Space Traversal
by: Pegios, Paraskevas, et al.
Published: (2024)
by: Pegios, Paraskevas, et al.
Published: (2024)
Attribution Explanations for Deep Neural Networks: A Theoretical Perspective
by: Deng, Huiqi, et al.
Published: (2025)
by: Deng, Huiqi, et al.
Published: (2025)
Explainable AI needs formalization
by: Haufe, Stefan, et al.
Published: (2024)
by: Haufe, Stefan, et al.
Published: (2024)
Controlling False Positives in Image Segmentation via Conformal Prediction
by: Mossina, Luca, et al.
Published: (2025)
by: Mossina, Luca, et al.
Published: (2025)
Danoliteracy of Generative Large Language Models
by: Holm, Søren Vejlgaard, et al.
Published: (2024)
by: Holm, Søren Vejlgaard, et al.
Published: (2024)
Connecting Concept Convexity and Human-Machine Alignment in Deep Neural Networks
by: Dorszewski, Teresa, et al.
Published: (2024)
by: Dorszewski, Teresa, et al.
Published: (2024)
BiSSL: Enhancing the Alignment Between Self-Supervised Pretraining and Downstream Fine-Tuning via Bilevel Optimization
by: Zakarias, Gustav Wagner, et al.
Published: (2024)
by: Zakarias, Gustav Wagner, et al.
Published: (2024)
FAME: Formal Abstract Minimal Explanation for Neural Networks
by: Boumazouza, Ryma, et al.
Published: (2026)
by: Boumazouza, Ryma, et al.
Published: (2026)
Localizing Anomalies in Critical Infrastructure using Model-Based Drift Explanations
by: Vaquet, Valerie, et al.
Published: (2023)
by: Vaquet, Valerie, et al.
Published: (2023)
Linear Explanations for Individual Neurons
by: Oikarinen, Tuomas, et al.
Published: (2024)
by: Oikarinen, Tuomas, et al.
Published: (2024)
Benchmarking Recursive-Collapse Warning Claims Under Matched False-Positive Control
by: Mullett, David
Published: (2026)
by: Mullett, David
Published: (2026)
Similar Items
-
Reasoning Models Don't Just Think Longer, They Move Differently
by: Gjølbye, Anders, et al.
Published: (2026) -
Missing-Data-Induced Phase Transitions in Spectral PLS for Multimodal Learning
by: Gjølbye, Anders, et al.
Published: (2026) -
Concept-based explainability for an EEG transformer model
by: Gjølbye, Anders, et al.
Published: (2023) -
Robustness of Visual Explanations to Common Data Augmentation
by: Tětková, Lenka, et al.
Published: (2023) -
cc-Shapley: Measuring Multivariate Feature Importance Needs Causal Context
by: Martin, Jörg, et al.
Published: (2026)