Prototypical Self-Explainable Models Without Re-training
Fuente:
arXiv
Guardado en:
| Autores principales: | Gautam, Srishti, Boubekki, Ahcene, Höhne, Marina M. C., Kampffmeyer, Michael C. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Pantypes: Diverse Representatives for Self-Explainable Models
por: Kjærsgaard, Rune, et al.
Publicado: (2024)
por: Kjærsgaard, Rune, et al.
Publicado: (2024)
Explainable AI needs formalization
por: Haufe, Stefan, et al.
Publicado: (2024)
por: Haufe, Stefan, et al.
Publicado: (2024)
Uncertainty Gating for Cost-Aware Explainable Artificial Intelligence
por: Mikriukov, Georgii, et al.
Publicado: (2026)
por: Mikriukov, Georgii, et al.
Publicado: (2026)
Supercm: Revisiting Clustering for Semi-Supervised Learning
por: Singh, Durgesh, et al.
Publicado: (2025)
por: Singh, Durgesh, et al.
Publicado: (2025)
REPEAT: Improving Uncertainty Estimation in Representation Learning Explainability
por: Wickstrøm, Kristoffer K., et al.
Publicado: (2024)
por: Wickstrøm, Kristoffer K., et al.
Publicado: (2024)
Finding the right XAI method -- A Guide for the Evaluation and Ranking of Explainable AI Methods in Climate Science
por: Bommer, Philine, et al.
Publicado: (2023)
por: Bommer, Philine, et al.
Publicado: (2023)
Post-hoc Self-explanation of CNNs
por: Boubekki, Ahcène, et al.
Publicado: (2026)
por: Boubekki, Ahcène, et al.
Publicado: (2026)
Labeling Neural Representations with Inverse Recognition
por: Bykov, Kirill, et al.
Publicado: (2023)
por: Bykov, Kirill, et al.
Publicado: (2023)
Sanity Checks Revisited: An Exploration to Repair the Model Parameter Randomisation Test
por: Hedström, Anna, et al.
Publicado: (2024)
por: Hedström, Anna, et al.
Publicado: (2024)
Deep Learning Meets Teleconnections: Improving S2S Predictions for European Winter Weather
por: Bommer, Philine L., et al.
Publicado: (2025)
por: Bommer, Philine L., et al.
Publicado: (2025)
Semantic Prototypes: Enhancing Transparency Without Black Boxes
por: Menis-Mastromichalakis, Orfeas, et al.
Publicado: (2024)
por: Menis-Mastromichalakis, Orfeas, et al.
Publicado: (2024)
Explaining the Impact of Training on Vision Models via Activation Clustering
por: Boubekki, Ahcène, et al.
Publicado: (2024)
por: Boubekki, Ahcène, et al.
Publicado: (2024)
Cross-domain Random Pre-training with Prototypes for Reinforcement Learning
por: Liu, Xin, et al.
Publicado: (2023)
por: Liu, Xin, et al.
Publicado: (2023)
CoSy: Evaluating Textual Explanations of Neurons
por: Kopf, Laura, et al.
Publicado: (2024)
por: Kopf, Laura, et al.
Publicado: (2024)
EXACT: Towards a platform for empirically benchmarking Machine Learning model explanation methods
por: Clark, Benedict, et al.
Publicado: (2024)
por: Clark, Benedict, et al.
Publicado: (2024)
A Fresh Look at Sanity Checks for Saliency Maps
por: Hedström, Anna, et al.
Publicado: (2024)
por: Hedström, Anna, et al.
Publicado: (2024)
Dynamic Policy Fusion for User Alignment Without Re-Interaction
por: Palattuparambil, Ajsal Shereef, et al.
Publicado: (2024)
por: Palattuparambil, Ajsal Shereef, et al.
Publicado: (2024)
Capturing Polysemanticity with PRISM: A Multi-Concept Feature Description Framework
por: Kopf, Laura, et al.
Publicado: (2025)
por: Kopf, Laura, et al.
Publicado: (2025)
SuperCM: Improving Semi-Supervised Learning and Domain Adaptation through differentiable clustering
por: Singh, Durgesh, et al.
Publicado: (2025)
por: Singh, Durgesh, et al.
Publicado: (2025)
Explainable AI for Fair Sepsis Mortality Predictive Model
por: Chang, Chia-Hsuan, et al.
Publicado: (2024)
por: Chang, Chia-Hsuan, et al.
Publicado: (2024)
Disentangled and Self-Explainable Node Representation Learning
por: Piaggesi, Simone, et al.
Publicado: (2024)
por: Piaggesi, Simone, et al.
Publicado: (2024)
Explainability of Machine Learning Models under Missing Data
por: Vo, Tuan L., et al.
Publicado: (2024)
por: Vo, Tuan L., et al.
Publicado: (2024)
Can We Predict Your Next Move Without Breaking Your Privacy?
por: Soni, Arpita, et al.
Publicado: (2025)
por: Soni, Arpita, et al.
Publicado: (2025)
SpecReX: Explainable AI for Raman Spectroscopy
por: Blake, Nathan, et al.
Publicado: (2025)
por: Blake, Nathan, et al.
Publicado: (2025)
Reconsidering Faithfulness in Regular, Self-Explainable and Domain Invariant GNNs
por: Azzolin, Steve, et al.
Publicado: (2024)
por: Azzolin, Steve, et al.
Publicado: (2024)
Explaining Bayesian Neural Networks
por: Bykov, Kirill, et al.
Publicado: (2021)
por: Bykov, Kirill, et al.
Publicado: (2021)
Self-ReSET: Learning to Self-Recover from Unsafe Reasoning Trajectories
por: Zhang, Dongcheng, et al.
Publicado: (2026)
por: Zhang, Dongcheng, et al.
Publicado: (2026)
Fixed Point Explainability
por: La Malfa, Emanuele, et al.
Publicado: (2025)
por: La Malfa, Emanuele, et al.
Publicado: (2025)
PFML: Self-Supervised Learning of Time-Series Data Without Representation Collapse
por: Vaaras, Einari, et al.
Publicado: (2024)
por: Vaaras, Einari, et al.
Publicado: (2024)
Discrete Prototypical Memories for Federated Time Series Foundation Models
por: Deng, Liwei, et al.
Publicado: (2026)
por: Deng, Liwei, et al.
Publicado: (2026)
Explaining AI Without Code: A User Study on Explainable AI
por: Abarca, Natalia, et al.
Publicado: (2025)
por: Abarca, Natalia, et al.
Publicado: (2025)
XAI-Units: Benchmarking Explainability Methods with Unit Tests
por: Lee, Jun Rui, et al.
Publicado: (2025)
por: Lee, Jun Rui, et al.
Publicado: (2025)
BeamVQ: Aligning Space-Time Forecasting Model via Self-training on Physics-aware Metrics
por: Wu, Hao, et al.
Publicado: (2024)
por: Wu, Hao, et al.
Publicado: (2024)
On the Limits of Self-Improving in Large Language Models: The Singularity Is Not Near Without Symbolic Model Synthesis
por: Zenil, Hector
Publicado: (2026)
por: Zenil, Hector
Publicado: (2026)
Value Imprint: A Technique for Auditing the Human Values Embedded in RLHF Datasets
por: Obi, Ike, et al.
Publicado: (2024)
por: Obi, Ike, et al.
Publicado: (2024)
Manipulating Feature Visualizations with Gradient Slingshots
por: Bareeva, Dilyara, et al.
Publicado: (2024)
por: Bareeva, Dilyara, et al.
Publicado: (2024)
Self-Exploring Language Models for Explainable Link Forecasting on Temporal Graphs via Reinforcement Learning
por: Ding, Zifeng, et al.
Publicado: (2025)
por: Ding, Zifeng, et al.
Publicado: (2025)
Can Multitask Learning Enhance Model Explainability?
por: Najjar, Hiba, et al.
Publicado: (2025)
por: Najjar, Hiba, et al.
Publicado: (2025)
Quantifying the Pre-training Dividend: Generative versus Latent Self-Supervised Learning for Time Series Foundation Models
por: Major, Noam, et al.
Publicado: (2026)
por: Major, Noam, et al.
Publicado: (2026)
Evaluating Large Language Models for Phishing Detection, Self-Consistency, Faithfulness, and Explainability
por: Kuikel, Shova, et al.
Publicado: (2025)
por: Kuikel, Shova, et al.
Publicado: (2025)
Ejemplares similares
-
Pantypes: Diverse Representatives for Self-Explainable Models
por: Kjærsgaard, Rune, et al.
Publicado: (2024) -
Explainable AI needs formalization
por: Haufe, Stefan, et al.
Publicado: (2024) -
Uncertainty Gating for Cost-Aware Explainable Artificial Intelligence
por: Mikriukov, Georgii, et al.
Publicado: (2026) -
Supercm: Revisiting Clustering for Semi-Supervised Learning
por: Singh, Durgesh, et al.
Publicado: (2025) -
REPEAT: Improving Uncertainty Estimation in Representation Learning Explainability
por: Wickstrøm, Kristoffer K., et al.
Publicado: (2024)