$α$-TCAV: A Unified Framework for Testing with Concept Activation Vectors
Fuente:
arXiv
Saved in:
| Main Authors: | Schnoor, Ekkehard, Said, Jawher, Tiomoko, Malik, Samek, Wojciech, Jung, Alexander |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Concept activation vectors: a unifying view and adversarial attacks
by: Schnoor, Ekkehard, et al.
Published: (2025)
by: Schnoor, Ekkehard, et al.
Published: (2025)
Incorporating priors in learning: a random matrix study under a teacher-student framework
by: Tiomoko, Malik, et al.
Published: (2025)
by: Tiomoko, Malik, et al.
Published: (2025)
E-TCAV: Formalizing Penultimate Proxies for Efficient Concept Based Interpretability
by: Aslam, Hasib, et al.
Published: (2026)
by: Aslam, Hasib, et al.
Published: (2026)
Navigating Neural Space: Revisiting Concept Activation Vectors to Overcome Directional Divergence
by: Pahde, Frederik, et al.
Published: (2022)
by: Pahde, Frederik, et al.
Published: (2022)
Nonparametric Identification of Latent Concepts
by: Zheng, Yujia, et al.
Published: (2025)
by: Zheng, Yujia, et al.
Published: (2025)
A Unified Theory of $θ$-Expectations
by: Qi, Qian
Published: (2025)
by: Qi, Qian
Published: (2025)
Model Science: getting serious about verification, explanation and control of AI systems
by: Biecek, Przemyslaw, et al.
Published: (2025)
by: Biecek, Przemyslaw, et al.
Published: (2025)
Post-Hoc Concept Disentanglement: From Correlated to Isolated Concept Representations
by: Erogullari, Eren, et al.
Published: (2025)
by: Erogullari, Eren, et al.
Published: (2025)
Position: Explain to Question not to Justify
by: Biecek, Przemyslaw, et al.
Published: (2024)
by: Biecek, Przemyslaw, et al.
Published: (2024)
High-Dimensional Analysis of Bootstrap Ensemble Classifiers
by: Tiomoko, Malik, et al.
Published: (2025)
by: Tiomoko, Malik, et al.
Published: (2025)
Iterative Inference in a Chess-Playing Neural Network
by: Sandmann, Elias, et al.
Published: (2025)
by: Sandmann, Elias, et al.
Published: (2025)
Visual-TCAV: Concept-based Attribution and Saliency Maps for Post-hoc Explainability in Image Classification
by: De Santis, Antonio, et al.
Published: (2024)
by: De Santis, Antonio, et al.
Published: (2024)
From Attribution Maps to Human-Understandable Explanations through Concept Relevance Propagation
by: Achtibat, Reduan, et al.
Published: (2022)
by: Achtibat, Reduan, et al.
Published: (2022)
Circuit Insights: Towards Interpretability Beyond Activations
by: Golimblevskaia, Elena, et al.
Published: (2025)
by: Golimblevskaia, Elena, et al.
Published: (2025)
Beyond Scalars: Concept-Based Alignment Analysis in Vision Transformers
by: Vielhaben, Johanna, et al.
Published: (2024)
by: Vielhaben, Johanna, et al.
Published: (2024)
Bayesian Inference with Deep Weakly Nonlinear Networks
by: Hanin, Boris, et al.
Published: (2024)
by: Hanin, Boris, et al.
Published: (2024)
Explaining Predictive Uncertainty by Exposing Second-Order Effects
by: Bley, Florian, et al.
Published: (2024)
by: Bley, Florian, et al.
Published: (2024)
A Mean-Field Theory of $Θ$-Expectations
by: Qi, Qian
Published: (2025)
by: Qi, Qian
Published: (2025)
Generalisation of Total Uncertainty in AI: A Theoretical Study
by: Shariatmadar, Keivan
Published: (2024)
by: Shariatmadar, Keivan
Published: (2024)
Greedy Selection under Independent Increments: A Toy Model Analysis
by: Yang, Huitao
Published: (2025)
by: Yang, Huitao
Published: (2025)
Beyond Propagation of Chaos: A Stochastic Algorithm for Mean Field Optimization
by: Tankala, Chandan, et al.
Published: (2025)
by: Tankala, Chandan, et al.
Published: (2025)
Representative Arm Identification: A fixed confidence approach to identify cluster representatives
by: Gharat, Sarvesh, et al.
Published: (2024)
by: Gharat, Sarvesh, et al.
Published: (2024)
Advancing Deep Learning through Probability Engineering: A Pragmatic Paradigm for Modern AI
by: Zhang, Jianyi
Published: (2025)
by: Zhang, Jianyi
Published: (2025)
On Training-Test (Mis)alignment in Unsupervised Combinatorial Optimization: Observation, Empirical Exploration, and Analysis
by: Bu, Fanchen, et al.
Published: (2025)
by: Bu, Fanchen, et al.
Published: (2025)
Note on Martingale Theory and Applications
by: Zou, Xiandong
Published: (2026)
by: Zou, Xiandong
Published: (2026)
Soft-to-Hard Routing in Sparse Mixture-of-Experts Models
by: Rastegar, Reza
Published: (2026)
by: Rastegar, Reza
Published: (2026)
On the Equivalence of Random Network Distillation, Deep Ensembles, and Bayesian Inference
by: Zanger, Moritz A., et al.
Published: (2026)
by: Zanger, Moritz A., et al.
Published: (2026)
Neural Network Parameter-optimization of Gaussian pmDAGs
by: Saremi, Mehrzad
Published: (2023)
by: Saremi, Mehrzad
Published: (2023)
Deep Conditional Measure Quantization
by: Turinici, Gabriel
Published: (2023)
by: Turinici, Gabriel
Published: (2023)
Feature Learning Dynamics in Infinite-Depth Neural Networks
by: Yao, Zihan, et al.
Published: (2025)
by: Yao, Zihan, et al.
Published: (2025)
Causal Effect Identification in Heterogeneous Environments from Higher-Order Moments
by: Kivva, Yaroslav, et al.
Published: (2025)
by: Kivva, Yaroslav, et al.
Published: (2025)
Explicit Density Approximation for Neural Implicit Samplers Using a Bernstein-Based Convex Divergence
by: de Frutos, José Manuel, et al.
Published: (2025)
by: de Frutos, José Manuel, et al.
Published: (2025)
Neural Expectation Operators
by: Qi, Qian
Published: (2025)
by: Qi, Qian
Published: (2025)
Efficient Training of Neural SDEs Using Stochastic Optimal Control
by: Daems, Rembert, et al.
Published: (2025)
by: Daems, Rembert, et al.
Published: (2025)
Generalized Gaussian Temporal Difference Error for Uncertainty-aware Reinforcement Learning
by: Kim, Seyeon, et al.
Published: (2024)
by: Kim, Seyeon, et al.
Published: (2024)
Quantitative CLTs in Deep Neural Networks
by: Favaro, Stefano, et al.
Published: (2023)
by: Favaro, Stefano, et al.
Published: (2023)
Neural Laplace for learning Stochastic Differential Equations
by: Carrel, Adrien
Published: (2024)
by: Carrel, Adrien
Published: (2024)
Synthetic Datasets for Machine Learning on Spatio-Temporal Graphs using PDEs
by: Arndt, Jost, et al.
Published: (2025)
by: Arndt, Jost, et al.
Published: (2025)
ECQ$^{\text{x}}$: Explainability-Driven Quantization for Low-Bit and Sparse DNNs
by: Becking, Daniel, et al.
Published: (2021)
by: Becking, Daniel, et al.
Published: (2021)
Atlas-Alignment: Making Interpretability Transferable Across Language Models
by: Puri, Bruno, et al.
Published: (2025)
by: Puri, Bruno, et al.
Published: (2025)
Similar Items
-
Concept activation vectors: a unifying view and adversarial attacks
by: Schnoor, Ekkehard, et al.
Published: (2025) -
Incorporating priors in learning: a random matrix study under a teacher-student framework
by: Tiomoko, Malik, et al.
Published: (2025) -
E-TCAV: Formalizing Penultimate Proxies for Efficient Concept Based Interpretability
by: Aslam, Hasib, et al.
Published: (2026) -
Navigating Neural Space: Revisiting Concept Activation Vectors to Overcome Directional Divergence
by: Pahde, Frederik, et al.
Published: (2022) -
Nonparametric Identification of Latent Concepts
by: Zheng, Yujia, et al.
Published: (2025)