Saved in:
| Main Authors: | Liu, Junhao, Yu, Haonan, Zhang, Xin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2410.12439 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MAnchors: Memorization-Based Acceleration of Anchors via Rule Reuse and Transformation
by: Yu, Haonan, et al.
Published: (2025)
by: Yu, Haonan, et al.
Published: (2025)
ReX: A Framework for Incorporating Temporal Information in Model-Agnostic Local Explanation Techniques
by: Liu, Junhao, et al.
Published: (2022)
by: Liu, Junhao, et al.
Published: (2022)
Revitalizing Black-Box Interpretability: Actionable Interpretability for LLMs via Proxy Models
by: Liu, Junhao, et al.
Published: (2025)
by: Liu, Junhao, et al.
Published: (2025)
Focus-LIME: Surgical Interpretation of Long-Context Large Language Models via Proxy-Based Neighborhood Selection
by: Liu, Junhao, et al.
Published: (2026)
by: Liu, Junhao, et al.
Published: (2026)
Unifying Attribution-Based Explanations Using Functional Decomposition
by: Gevaert, Arne, et al.
Published: (2024)
by: Gevaert, Arne, et al.
Published: (2024)
From Attribution Maps to Human-Understandable Explanations through Concept Relevance Propagation
by: Achtibat, Reduan, et al.
Published: (2022)
by: Achtibat, Reduan, et al.
Published: (2022)
Beyond the Academic Monoculture: A Unified Framework and Industrial Perspective for Attributed Graph Clustering
by: Liu, Yunhui, et al.
Published: (2026)
by: Liu, Yunhui, et al.
Published: (2026)
Missingness Bias Calibration in Feature Attribution Explanations
by: Sridhar, Shailesh, et al.
Published: (2026)
by: Sridhar, Shailesh, et al.
Published: (2026)
Counterfactual Explanations Under Concept Drift
by: Kostrzewa, Marcin, et al.
Published: (2026)
by: Kostrzewa, Marcin, et al.
Published: (2026)
An Axiomatic Approach to Model-Agnostic Concept Explanations
by: Feng, Zhili, et al.
Published: (2024)
by: Feng, Zhili, et al.
Published: (2024)
Attribution Explanations for Deep Neural Networks: A Theoretical Perspective
by: Deng, Huiqi, et al.
Published: (2025)
by: Deng, Huiqi, et al.
Published: (2025)
Unifying Perspectives: Plausible Counterfactual Explanations on Global, Group-wise, and Local Levels
by: Furman, Oleksii, et al.
Published: (2024)
by: Furman, Oleksii, et al.
Published: (2024)
Beyond Linear Steering: Unified Multi-Attribute Control for Language Models
by: Oozeer, Narmeen, et al.
Published: (2025)
by: Oozeer, Narmeen, et al.
Published: (2025)
Explaining Concept Shift with Interpretable Feature Attribution
by: Lyu, Ruiqi, et al.
Published: (2025)
by: Lyu, Ruiqi, et al.
Published: (2025)
Backward Compatibility in Attributive Explanation and Enhanced Model Training Method
by: Matsuno, Ryuta
Published: (2024)
by: Matsuno, Ryuta
Published: (2024)
Minimizing False-Positive Attributions in Explanations of Non-Linear Models
by: Gjølbye, Anders, et al.
Published: (2025)
by: Gjølbye, Anders, et al.
Published: (2025)
Direct Preference Optimization for Adaptive Concept-based Explanations
by: Teneggi, Jacopo, et al.
Published: (2025)
by: Teneggi, Jacopo, et al.
Published: (2025)
Self-explaining Neural Network with Concept-based Explanations for ICU Mortality Prediction
by: Kumar, Sayantan, et al.
Published: (2021)
by: Kumar, Sayantan, et al.
Published: (2021)
Exploring Concept Subspace for Self-explainable Text-Attributed Graph Learning
by: Han, Xiaoxue, et al.
Published: (2026)
by: Han, Xiaoxue, et al.
Published: (2026)
Global Concept Explanations for Graphs by Contrastive Learning
by: Teufel, Jonas, et al.
Published: (2024)
by: Teufel, Jonas, et al.
Published: (2024)
Neuron-Level Knowledge Attribution in Large Language Models
by: Yu, Zeping, et al.
Published: (2023)
by: Yu, Zeping, et al.
Published: (2023)
When Explanations Lie: Why Many Modified BP Attributions Fail
by: Sixt, Leon, et al.
Published: (2019)
by: Sixt, Leon, et al.
Published: (2019)
Causal Explanation of Concept Drift -- A Truly Actionable Approach
by: Komnick, David, et al.
Published: (2025)
by: Komnick, David, et al.
Published: (2025)
Subgraph Concept Networks: Concept Levels in Graph Classification
by: Magister, Lucie Charlotte, et al.
Published: (2026)
by: Magister, Lucie Charlotte, et al.
Published: (2026)
Can LLMs Learn New Concepts Incrementally without Forgetting?
by: Zheng, Junhao, et al.
Published: (2024)
by: Zheng, Junhao, et al.
Published: (2024)
Unisoma: A Unified Transformer-based Solver for Multi-Solid Systems
by: Tao, Shilong, et al.
Published: (2025)
by: Tao, Shilong, et al.
Published: (2025)
Learning Concept Bottleneck Models from Mechanistic Explanations
by: De Santis, Antonio, et al.
Published: (2026)
by: De Santis, Antonio, et al.
Published: (2026)
Unveiling Concept Attribution in Diffusion Models
by: Nguyen, Quang H., et al.
Published: (2024)
by: Nguyen, Quang H., et al.
Published: (2024)
Model-Based Counterfactual Explanations Incorporating Feature Space Attributes for Tabular Data
by: Sumiya, Yuta, et al.
Published: (2024)
by: Sumiya, Yuta, et al.
Published: (2024)
ConceptFlow: Hierarchical and Fine-grained Concept-Based Explanation for Convolutional Neural Networks
by: Mu, Xinyu, et al.
Published: (2025)
by: Mu, Xinyu, et al.
Published: (2025)
Explaining Hypergraph Neural Networks: From Local Explanations to Global Concepts
by: Su, Shiye, et al.
Published: (2024)
by: Su, Shiye, et al.
Published: (2024)
Estimation of Concept Explanations Should be Uncertainty Aware
by: Piratla, Vihari, et al.
Published: (2023)
by: Piratla, Vihari, et al.
Published: (2023)
Disrupting Model Merging: A Parameter-Level Defense Without Sacrificing Accuracy
by: Junhao, Wei, et al.
Published: (2025)
by: Junhao, Wei, et al.
Published: (2025)
Towards a Unified Framework for Evaluating Explanations
by: Pinto, Juan D., et al.
Published: (2024)
by: Pinto, Juan D., et al.
Published: (2024)
Harnessing Explanations: LLM-to-LM Interpreter for Enhanced Text-Attributed Graph Representation Learning
by: He, Xiaoxin, et al.
Published: (2023)
by: He, Xiaoxin, et al.
Published: (2023)
Concept-Based Abductive and Contrastive Explanations for Behaviors of Vision Models
by: Canizales, Ronaldo, et al.
Published: (2026)
by: Canizales, Ronaldo, et al.
Published: (2026)
GCFX: Generative Counterfactual Explanations for Deep Graph Models at the Model Level
by: Hu, Jinlong, et al.
Published: (2026)
by: Hu, Jinlong, et al.
Published: (2026)
Unified Explanations in Machine Learning Models: A Perturbation Approach
by: Dineen, Jacob, et al.
Published: (2024)
by: Dineen, Jacob, et al.
Published: (2024)
Evaluating Neuron Explanations: A Unified Framework with Sanity Checks
by: Oikarinen, Tuomas, et al.
Published: (2025)
by: Oikarinen, Tuomas, et al.
Published: (2025)
CohEx: A Generalized Framework for Cohort Explanation
by: Meng, Fanyu, et al.
Published: (2024)
by: Meng, Fanyu, et al.
Published: (2024)
Similar Items
-
MAnchors: Memorization-Based Acceleration of Anchors via Rule Reuse and Transformation
by: Yu, Haonan, et al.
Published: (2025) -
ReX: A Framework for Incorporating Temporal Information in Model-Agnostic Local Explanation Techniques
by: Liu, Junhao, et al.
Published: (2022) -
Revitalizing Black-Box Interpretability: Actionable Interpretability for LLMs via Proxy Models
by: Liu, Junhao, et al.
Published: (2025) -
Focus-LIME: Surgical Interpretation of Long-Context Large Language Models via Proxy-Based Neighborhood Selection
by: Liu, Junhao, et al.
Published: (2026) -
Unifying Attribution-Based Explanations Using Functional Decomposition
by: Gevaert, Arne, et al.
Published: (2024)