The Clever Hans Effect in Unsupervised Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Kauffmann, Jacob, Dippel, Jonas, Ruff, Lukas, Samek, Wojciech, Müller, Klaus-Robert, Montavon, Grégoire |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reliable Modeling of Distribution Shifts via Displacement-Reshaped Optimal Transport
by: Naumann, Philip, et al.
Published: (2026)
by: Naumann, Philip, et al.
Published: (2026)
Explaining Predictive Uncertainty by Exposing Second-Order Effects
by: Bley, Florian, et al.
Published: (2024)
by: Bley, Florian, et al.
Published: (2024)
Fast and Accurate Explanations of Distance-Based Classifiers by Uncovering Latent Explanatory Structures
by: Bley, Florian, et al.
Published: (2025)
by: Bley, Florian, et al.
Published: (2025)
Wasserstein Distances Made Explainable: Insights Into Dataset Shifts and Transport Phenomena
by: Naumann, Philip, et al.
Published: (2025)
by: Naumann, Philip, et al.
Published: (2025)
Mitigating Clever Hans Strategies in Image Classifiers through Generating Counterexamples
by: Bender, Sidney, et al.
Published: (2025)
by: Bender, Sidney, et al.
Published: (2025)
MambaLRP: Explaining Selective State Space Sequence Models
by: Jafari, Farnoush Rezaei, et al.
Published: (2024)
by: Jafari, Farnoush Rezaei, et al.
Published: (2024)
Uncovering the Structure of Explanation Quality with Spectral Analysis
by: Maeß, Johannes, et al.
Published: (2025)
by: Maeß, Johannes, et al.
Published: (2025)
Disentangled Explanations of Neural Network Predictions by Finding Relevant Subspaces
by: Chormai, Pattarawat, et al.
Published: (2022)
by: Chormai, Pattarawat, et al.
Published: (2022)
Towards Symbolic XAI -- Explanation Through Human Understandable Logical Relationships Between Features
by: Schnake, Thomas, et al.
Published: (2024)
by: Schnake, Thomas, et al.
Published: (2024)
Model Science: getting serious about verification, explanation and control of AI systems
by: Biecek, Przemyslaw, et al.
Published: (2025)
by: Biecek, Przemyslaw, et al.
Published: (2025)
Insightful analysis of historical sources at scales beyond human capabilities using unsupervised Machine Learning and XAI
by: Eberle, Oliver, et al.
Published: (2023)
by: Eberle, Oliver, et al.
Published: (2023)
Uncertainty Gating for Cost-Aware Explainable Artificial Intelligence
by: Mikriukov, Georgii, et al.
Published: (2026)
by: Mikriukov, Georgii, et al.
Published: (2026)
Position: Explain to Question not to Justify
by: Biecek, Przemyslaw, et al.
Published: (2024)
by: Biecek, Przemyslaw, et al.
Published: (2024)
Iterative Inference in a Chess-Playing Neural Network
by: Sandmann, Elias, et al.
Published: (2025)
by: Sandmann, Elias, et al.
Published: (2025)
Optimizing Federated Learning by Entropy-Based Client Selection
by: Lutz, Andreas, et al.
Published: (2024)
by: Lutz, Andreas, et al.
Published: (2024)
ECQ$^{\text{x}}$: Explainability-Driven Quantization for Low-Bit and Sparse DNNs
by: Becking, Daniel, et al.
Published: (2021)
by: Becking, Daniel, et al.
Published: (2021)
MeDi: Metadata-Guided Diffusion Models for Mitigating Biases in Tumor Classification
by: Drexlin, David Jacob, et al.
Published: (2025)
by: Drexlin, David Jacob, et al.
Published: (2025)
Synthetic Datasets for Machine Learning on Spatio-Temporal Graphs using PDEs
by: Arndt, Jost, et al.
Published: (2025)
by: Arndt, Jost, et al.
Published: (2025)
XpertAI: uncovering regression model strategies for sub-manifolds
by: Letzgus, Simon, et al.
Published: (2024)
by: Letzgus, Simon, et al.
Published: (2024)
Investigating the Robustness of Subtask Distillation under Spurious Correlation
by: Chormai, Pattarawat, et al.
Published: (2026)
by: Chormai, Pattarawat, et al.
Published: (2026)
Opportunities and limitations of explaining quantum machine learning
by: Gil-Fuster, Elies, et al.
Published: (2024)
by: Gil-Fuster, Elies, et al.
Published: (2024)
Heuristic Transformer: Belief Augmented In-Context Reinforcement Learning
by: Dippel, Oliver, et al.
Published: (2025)
by: Dippel, Oliver, et al.
Published: (2025)
Reproducibility study on how to find Spurious Correlations, Shortcut Learning, Clever Hans or Group-Distributional non-robustness and how to fix them
by: Delzer, Ole, et al.
Published: (2026)
by: Delzer, Ole, et al.
Published: (2026)
Atlas-Alignment: Making Interpretability Transferable Across Language Models
by: Puri, Bruno, et al.
Published: (2025)
by: Puri, Bruno, et al.
Published: (2025)
Building Trust in PINNs: Error Estimation through Finite Difference Methods
by: Krasowski, Aleksander, et al.
Published: (2026)
by: Krasowski, Aleksander, et al.
Published: (2026)
CleverCatch: A Knowledge-Guided Weak Supervision Model for Fraud Detection
by: Mozafari, Amirhossein, et al.
Published: (2025)
by: Mozafari, Amirhossein, et al.
Published: (2025)
Sparse, Efficient and Explainable Data Attribution with DualXDA
by: Yolcu, Galip Ümit, et al.
Published: (2024)
by: Yolcu, Galip Ümit, et al.
Published: (2024)
$α$-TCAV: A Unified Framework for Testing with Concept Activation Vectors
by: Schnoor, Ekkehard, et al.
Published: (2026)
by: Schnoor, Ekkehard, et al.
Published: (2026)
Human alignment of neural network representations
by: Muttenthaler, Lukas, et al.
Published: (2022)
by: Muttenthaler, Lukas, et al.
Published: (2022)
From What to How: Attributing CLIP's Latent Components Reveals Unexpected Semantic Reliance
by: Dreyer, Maximilian, et al.
Published: (2025)
by: Dreyer, Maximilian, et al.
Published: (2025)
Towards Robust Foundation Models for Digital Pathology
by: Kömen, Jonah, et al.
Published: (2025)
by: Kömen, Jonah, et al.
Published: (2025)
Distilling Lightweight Domain Experts from Large ML Models by Identifying Relevant Subspaces
by: Chormai, Pattarawat, et al.
Published: (2026)
by: Chormai, Pattarawat, et al.
Published: (2026)
Knowledge-Free Correlated Agreement for Incentivizing Federated Learning
by: Witt, Leon, et al.
Published: (2026)
by: Witt, Leon, et al.
Published: (2026)
Post-Hoc Concept Disentanglement: From Correlated to Isolated Concept Representations
by: Erogullari, Eren, et al.
Published: (2025)
by: Erogullari, Eren, et al.
Published: (2025)
Ensuring Medical AI Safety: Interpretability-Driven Detection and Mitigation of Spurious Model Behavior and Associated Data
by: Pahde, Frederik, et al.
Published: (2025)
by: Pahde, Frederik, et al.
Published: (2025)
Objective drives the consistency of representational similarity across datasets
by: Ciernik, Laure, et al.
Published: (2024)
by: Ciernik, Laure, et al.
Published: (2024)
Mechanistic understanding and validation of large AI models with SemanticLens
by: Dreyer, Maximilian, et al.
Published: (2025)
by: Dreyer, Maximilian, et al.
Published: (2025)
From Attribution Maps to Human-Understandable Explanations through Concept Relevance Propagation
by: Achtibat, Reduan, et al.
Published: (2022)
by: Achtibat, Reduan, et al.
Published: (2022)
Relevance-driven Input Dropout: an Explanation-guided Regularization Technique
by: Gururaj, Shreyas, et al.
Published: (2025)
by: Gururaj, Shreyas, et al.
Published: (2025)
Quanda: An Interpretability Toolkit for Training Data Attribution Evaluation and Beyond
by: Bareeva, Dilyara, et al.
Published: (2024)
by: Bareeva, Dilyara, et al.
Published: (2024)
Similar Items
-
Reliable Modeling of Distribution Shifts via Displacement-Reshaped Optimal Transport
by: Naumann, Philip, et al.
Published: (2026) -
Explaining Predictive Uncertainty by Exposing Second-Order Effects
by: Bley, Florian, et al.
Published: (2024) -
Fast and Accurate Explanations of Distance-Based Classifiers by Uncovering Latent Explanatory Structures
by: Bley, Florian, et al.
Published: (2025) -
Wasserstein Distances Made Explainable: Insights Into Dataset Shifts and Transport Phenomena
by: Naumann, Philip, et al.
Published: (2025) -
Mitigating Clever Hans Strategies in Image Classifiers through Generating Counterexamples
by: Bender, Sidney, et al.
Published: (2025)