ABE: A Unified Framework for Robust and Faithful Attribution-Based Explainability
Fuente:
arXiv
Saved in:
| Main Authors: | Zhu, Zhiyu, Zhang, Jiayu, Jin, Zhibo, Chen, Fang, Zhou, Jianlong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Enhancing Model Interpretability with Local Attribution over Global Exploration
by: Zhu, Zhiyu, et al.
Published: (2024)
by: Zhu, Zhiyu, et al.
Published: (2024)
Attribution for Enhanced Explanation with Transferable Adversarial eXploration
by: Zhu, Zhiyu, et al.
Published: (2024)
by: Zhu, Zhiyu, et al.
Published: (2024)
Revisiting Sharpness-Aware Minimization: A More Faithful and Effective Implementation
by: Chen, Jianlong, et al.
Published: (2026)
by: Chen, Jianlong, et al.
Published: (2026)
Leveraging Information Consistency in Frequency and Spatial Domain for Adversarial Attacks
by: Jin, Zhibo, et al.
Published: (2024)
by: Jin, Zhibo, et al.
Published: (2024)
F-Fidelity: A Robust Framework for Faithfulness Evaluation of Explainable AI
by: Zheng, Xu, et al.
Published: (2024)
by: Zheng, Xu, et al.
Published: (2024)
Enhancing Transferability of Adversarial Attacks with GE-AdvGAN+: A Comprehensive Framework for Gradient Editing
by: Jin, Zhibo, et al.
Published: (2024)
by: Jin, Zhibo, et al.
Published: (2024)
PAR-AdvGAN: Improving Adversarial Attack Capability with Progressive Auto-Regression AdvGAN
by: Zhang, Jiayu, et al.
Published: (2025)
by: Zhang, Jiayu, et al.
Published: (2025)
Beyond Noisy-TVs: Noise-Robust Exploration Via Learning Progress Monitoring
by: Hou, Zhibo, et al.
Published: (2025)
by: Hou, Zhibo, et al.
Published: (2025)
Hybrid Attribution Priors for Explainable and Robust Model Training
by: Zhang, Zhuoran, et al.
Published: (2025)
by: Zhang, Zhuoran, et al.
Published: (2025)
DeepFaith: A Domain-Free and Model-Agnostic Unified Framework for Highly Faithful Explanations
by: Guo, Yuhan, et al.
Published: (2025)
by: Guo, Yuhan, et al.
Published: (2025)
Towards Unified Attribution in Explainable AI, Data-Centric AI, and Mechanistic Interpretability
by: Zhang, Shichang, et al.
Published: (2025)
by: Zhang, Shichang, et al.
Published: (2025)
Correlation-Aware Feature Attribution Based Explainable AI
by: Sengupta, Poushali, et al.
Published: (2025)
by: Sengupta, Poushali, et al.
Published: (2025)
Beyond Output Faithfulness: Learning Attributions that Preserve Computational Pathways
by: Zhang, Siyu, et al.
Published: (2025)
by: Zhang, Siyu, et al.
Published: (2025)
Explainability in Generative Medical Diffusion Models: A Faithfulness-Based Analysis on MRI Synthesis
by: Dey, Surjo, et al.
Published: (2026)
by: Dey, Surjo, et al.
Published: (2026)
Enhancing Adversarial Attacks via Parameter Adaptive Adversarial Attack
by: Jin, Zhibo, et al.
Published: (2024)
by: Jin, Zhibo, et al.
Published: (2024)
Optimal Look-back Horizon for Time Series Forecasting in Federated Learning
by: Tang, Dahao, et al.
Published: (2025)
by: Tang, Dahao, et al.
Published: (2025)
Narrowing Information Bottleneck Theory for Multimodal Image-Text Representations Interpretability
by: Zhu, Zhiyu, et al.
Published: (2025)
by: Zhu, Zhiyu, et al.
Published: (2025)
Reconsidering Faithfulness in Regular, Self-Explainable and Domain Invariant GNNs
by: Azzolin, Steve, et al.
Published: (2024)
by: Azzolin, Steve, et al.
Published: (2024)
Unifying VXAI: A Systematic Review and Framework for the Evaluation of Explainable AI
by: Dembinsky, David, et al.
Published: (2025)
by: Dembinsky, David, et al.
Published: (2025)
Incorporating Attribution Importance for Improving Faithfulness Metrics
by: Zhao, Zhixue, et al.
Published: (2023)
by: Zhao, Zhixue, et al.
Published: (2023)
Unifying Attribution-Based Explanations Using Functional Decomposition
by: Gevaert, Arne, et al.
Published: (2024)
by: Gevaert, Arne, et al.
Published: (2024)
Unraveling LoRA Interference: Orthogonal Subspaces for Robust Model Merging
by: Zhang, Haobo, et al.
Published: (2025)
by: Zhang, Haobo, et al.
Published: (2025)
AI Pangaea: Unifying Intelligence Islands for Adapting Myriad Tasks
by: Chang, Jianlong, et al.
Published: (2025)
by: Chang, Jianlong, et al.
Published: (2025)
Learning Unified Distance Metric for Heterogeneous Attribute Data Clustering
by: Zhang, Yiqun, et al.
Published: (2026)
by: Zhang, Yiqun, et al.
Published: (2026)
LEGO: A Lightweight and Efficient Multiple-Attribute Unlearning Framework for Recommender Systems
by: Yu, Fengyuan, et al.
Published: (2025)
by: Yu, Fengyuan, et al.
Published: (2025)
Enhancing Robustness of Offline Reinforcement Learning Under Data Corruption via Sharpness-Aware Minimization
by: Xu, Le, et al.
Published: (2025)
by: Xu, Le, et al.
Published: (2025)
Does Faithfulness Conflict with Plausibility? An Empirical Study in Explainable AI across NLP Tasks
by: Lu, Xiaolei, et al.
Published: (2024)
by: Lu, Xiaolei, et al.
Published: (2024)
Sparse, Efficient and Explainable Data Attribution with DualXDA
by: Yolcu, Galip Ümit, et al.
Published: (2024)
by: Yolcu, Galip Ümit, et al.
Published: (2024)
Towards a Unified Framework of Clustering-based Anomaly Detection
by: Fang, Zeyu, et al.
Published: (2024)
by: Fang, Zeyu, et al.
Published: (2024)
UFO: A Unified Flow-Oriented Framework for Robust Continual Graph Learning
by: Zhang, Danhui, et al.
Published: (2026)
by: Zhang, Danhui, et al.
Published: (2026)
A Unified Framework for Generative Data Augmentation: A Comprehensive Survey
by: Chen, Yunhao, et al.
Published: (2023)
by: Chen, Yunhao, et al.
Published: (2023)
High-order Knowledge Based Network Controllability Robustness Prediction: A Hypergraph Neural Network Approach
by: Mo, Shibing, et al.
Published: (2026)
by: Mo, Shibing, et al.
Published: (2026)
Explainable Human Activity Recognition: A Unified Review of Concepts and Mechanisms
by: Kundu, Mainak, et al.
Published: (2026)
by: Kundu, Mainak, et al.
Published: (2026)
FedEAT: A Robustness Optimization Framework for Federated LLMs
by: Pang, Yahao, et al.
Published: (2025)
by: Pang, Yahao, et al.
Published: (2025)
Faithfulness Evaluation for Decoder-only LLM Attributions with Controlled Retained Information
by: Huang, Xin, et al.
Published: (2026)
by: Huang, Xin, et al.
Published: (2026)
A Unified Frequency Domain Decomposition Framework for Interpretable and Robust Time Series Forecasting
by: He, Cheng, et al.
Published: (2025)
by: He, Cheng, et al.
Published: (2025)
Faithful Path Language Modeling for Explainable Recommendation over Knowledge Graph
by: Balloccu, Giacomo, et al.
Published: (2023)
by: Balloccu, Giacomo, et al.
Published: (2023)
Transformer Circuit Faithfulness Metrics are not Robust
by: Miller, Joseph, et al.
Published: (2024)
by: Miller, Joseph, et al.
Published: (2024)
Wrapper Boxes: Faithful Attribution of Model Predictions to Training Data
by: Su, Yiheng, et al.
Published: (2023)
by: Su, Yiheng, et al.
Published: (2023)
ANO: A Principled Approach to Robust Policy Optimization
by: Zhang, Yiheng, et al.
Published: (2026)
by: Zhang, Yiheng, et al.
Published: (2026)
Similar Items
-
Enhancing Model Interpretability with Local Attribution over Global Exploration
by: Zhu, Zhiyu, et al.
Published: (2024) -
Attribution for Enhanced Explanation with Transferable Adversarial eXploration
by: Zhu, Zhiyu, et al.
Published: (2024) -
Revisiting Sharpness-Aware Minimization: A More Faithful and Effective Implementation
by: Chen, Jianlong, et al.
Published: (2026) -
Leveraging Information Consistency in Frequency and Spatial Domain for Adversarial Attacks
by: Jin, Zhibo, et al.
Published: (2024) -
F-Fidelity: A Robust Framework for Faithfulness Evaluation of Explainable AI
by: Zheng, Xu, et al.
Published: (2024)