Enhancing Model Interpretability with Local Attribution over Global Exploration
Fuente:
arXiv
Saved in:
| Main Authors: | Zhu, Zhiyu, Jin, Zhibo, Zhang, Jiayu, Chen, Huaming |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ABE: A Unified Framework for Robust and Faithful Attribution-Based Explainability
by: Zhu, Zhiyu, et al.
Published: (2025)
by: Zhu, Zhiyu, et al.
Published: (2025)
Leveraging Information Consistency in Frequency and Spatial Domain for Adversarial Attacks
by: Jin, Zhibo, et al.
Published: (2024)
by: Jin, Zhibo, et al.
Published: (2024)
PAR-AdvGAN: Improving Adversarial Attack Capability with Progressive Auto-Regression AdvGAN
by: Zhang, Jiayu, et al.
Published: (2025)
by: Zhang, Jiayu, et al.
Published: (2025)
Attribution for Enhanced Explanation with Transferable Adversarial eXploration
by: Zhu, Zhiyu, et al.
Published: (2024)
by: Zhu, Zhiyu, et al.
Published: (2024)
Beyond Noisy-TVs: Noise-Robust Exploration Via Learning Progress Monitoring
by: Hou, Zhibo, et al.
Published: (2025)
by: Hou, Zhibo, et al.
Published: (2025)
Benchmarking Transferable Adversarial Attacks
by: Jin, Zhibo, et al.
Published: (2024)
by: Jin, Zhibo, et al.
Published: (2024)
DMS: Addressing Information Loss with More Steps for Pragmatic Adversarial Attacks
by: Zhu, Zhiyu, et al.
Published: (2024)
by: Zhu, Zhiyu, et al.
Published: (2024)
Optimal Look-back Horizon for Time Series Forecasting in Federated Learning
by: Tang, Dahao, et al.
Published: (2025)
by: Tang, Dahao, et al.
Published: (2025)
SL-CBM: Enhancing Concept Bottleneck Models with Semantic Locality for Better Interpretability
by: Zhang, Hanwei, et al.
Published: (2026)
by: Zhang, Hanwei, et al.
Published: (2026)
AIM: Attributing, Interpreting, Mitigating Data Unfairness
by: Liu, Zhining, et al.
Published: (2024)
by: Liu, Zhining, et al.
Published: (2024)
Learning to Price: Interpretable Attribute-Level Models for Dynamic Markets
by: Sethuraman, Srividhya, et al.
Published: (2026)
by: Sethuraman, Srividhya, et al.
Published: (2026)
Using a Local Surrogate Model to Interpret Temporal Shifts in Global Annual Data
by: Nakano, Shou, et al.
Published: (2024)
by: Nakano, Shou, et al.
Published: (2024)
What You See Is Not Always What You Get: Evaluating GPT's Comprehension of Source Code
by: Wen, Jiawen, et al.
Published: (2024)
by: Wen, Jiawen, et al.
Published: (2024)
ST-TGExplainer: Disentangling Stability and Transition Patterns for Temporal GNN Interpretability
by: Chen, Hongjiang, et al.
Published: (2026)
by: Chen, Hongjiang, et al.
Published: (2026)
Attributions All the Way Down? The Metagame of Interpretability
by: Baniecki, Hubert, et al.
Published: (2026)
by: Baniecki, Hubert, et al.
Published: (2026)
Enhancing Robustness of Offline Reinforcement Learning Under Data Corruption via Sharpness-Aware Minimization
by: Xu, Le, et al.
Published: (2025)
by: Xu, Le, et al.
Published: (2025)
FreqLens: Interpretable Frequency Attribution for Time Series Forecasting
by: Chen, Chi-Sheng, et al.
Published: (2026)
by: Chen, Chi-Sheng, et al.
Published: (2026)
GraphCLIP: Enhancing Transferability in Graph Foundation Models for Text-Attributed Graphs
by: Zhu, Yun, et al.
Published: (2024)
by: Zhu, Yun, et al.
Published: (2024)
VRAIL: Vectorized Reward-based Attribution for Interpretable Learning
by: Kim, Jina, et al.
Published: (2025)
by: Kim, Jina, et al.
Published: (2025)
GradCFA: A Hybrid Gradient-Based Counterfactual and Feature Attribution Explanation Algorithm for Local Interpretation of Neural Networks
by: Sanderson, Jacob, et al.
Published: (2026)
by: Sanderson, Jacob, et al.
Published: (2026)
Towards Unified Attribution in Explainable AI, Data-Centric AI, and Mechanistic Interpretability
by: Zhang, Shichang, et al.
Published: (2025)
by: Zhang, Shichang, et al.
Published: (2025)
EVGeoQA: Benchmarking LLMs on Dynamic, Multi-Objective Geo-Spatial Exploration
by: Wu, Jianfei, et al.
Published: (2026)
by: Wu, Jianfei, et al.
Published: (2026)
Quanda: An Interpretability Toolkit for Training Data Attribution Evaluation and Beyond
by: Bareeva, Dilyara, et al.
Published: (2024)
by: Bareeva, Dilyara, et al.
Published: (2024)
Causality-Aware Local Interpretable Model-Agnostic Explanations
by: Cinquini, Martina, et al.
Published: (2022)
by: Cinquini, Martina, et al.
Published: (2022)
Hybrid Attribution Priors for Explainable and Robust Model Training
by: Zhang, Zhuoran, et al.
Published: (2025)
by: Zhang, Zhuoran, et al.
Published: (2025)
FA-INR: Adaptive Implicit Neural Representations for Interpretable Exploration of Simulation Ensembles
by: Li, Ziwei, et al.
Published: (2025)
by: Li, Ziwei, et al.
Published: (2025)
Enhancing Model Privacy in Federated Learning with Random Masking and Quantization
by: Xu, Zhibo, et al.
Published: (2025)
by: Xu, Zhibo, et al.
Published: (2025)
Locally Interpretable Individualized Treatment Rules for Black-Box Decision Models
by: Charvadeh, Yasin Khadem, et al.
Published: (2026)
by: Charvadeh, Yasin Khadem, et al.
Published: (2026)
SHapley Estimated Explanation (SHEP): A Fast Post-Hoc Attribution Method for Interpreting Intelligent Fault Diagnosis
by: Chen, Qian, et al.
Published: (2025)
by: Chen, Qian, et al.
Published: (2025)
Feature-Selective Representation Misdirection for Machine Unlearning
by: Chen, Taozhao, et al.
Published: (2025)
by: Chen, Taozhao, et al.
Published: (2025)
Concept Influence: Leveraging Interpretability to Improve Performance and Efficiency in Training Data Attribution
by: Kowal, Matthew, et al.
Published: (2026)
by: Kowal, Matthew, et al.
Published: (2026)
Representational Homomorphism Predicts and Improves Compositional Generalization In Transformer Language Model
by: An, Zhiyu, et al.
Published: (2026)
by: An, Zhiyu, et al.
Published: (2026)
External Model Motivated Agents: Reinforcement Learning for Enhanced Environment Sampling
by: Bhagat, Rishav, et al.
Published: (2024)
by: Bhagat, Rishav, et al.
Published: (2024)
Mechanistic Data Attribution: Tracing the Training Origins of Interpretable LLM Units
by: Chen, Jianhui, et al.
Published: (2026)
by: Chen, Jianhui, et al.
Published: (2026)
Enhancing Transferability of Adversarial Attacks with GE-AdvGAN+: A Comprehensive Framework for Gradient Editing
by: Jin, Zhibo, et al.
Published: (2024)
by: Jin, Zhibo, et al.
Published: (2024)
KDSelector: A Knowledge-Enhanced and Data-Efficient Model Selector Learning Framework for Time Series Anomaly Detection
by: Liang, Zhiyu, et al.
Published: (2025)
by: Liang, Zhiyu, et al.
Published: (2025)
Scalable Hyperparameter-Divergent Ensemble Training with Automatic Learning Rate Exploration for Large Models
by: Cheng, Hailing, et al.
Published: (2026)
by: Cheng, Hailing, et al.
Published: (2026)
Interpretability-by-Design with Accurate Locally Additive Models and Conditional Feature Effects
by: Gkolemis, Vasilis, et al.
Published: (2026)
by: Gkolemis, Vasilis, et al.
Published: (2026)
Revisiting Data Attribution for Influence Functions
by: Zhu, Hongbo, et al.
Published: (2025)
by: Zhu, Hongbo, et al.
Published: (2025)
HiMAC: Hierarchical Macro-Micro Learning for Long-Horizon LLM Agents
by: Jin, Hongbo, et al.
Published: (2026)
by: Jin, Hongbo, et al.
Published: (2026)
Similar Items
-
ABE: A Unified Framework for Robust and Faithful Attribution-Based Explainability
by: Zhu, Zhiyu, et al.
Published: (2025) -
Leveraging Information Consistency in Frequency and Spatial Domain for Adversarial Attacks
by: Jin, Zhibo, et al.
Published: (2024) -
PAR-AdvGAN: Improving Adversarial Attack Capability with Progressive Auto-Regression AdvGAN
by: Zhang, Jiayu, et al.
Published: (2025) -
Attribution for Enhanced Explanation with Transferable Adversarial eXploration
by: Zhu, Zhiyu, et al.
Published: (2024) -
Beyond Noisy-TVs: Noise-Robust Exploration Via Learning Progress Monitoring
by: Hou, Zhibo, et al.
Published: (2025)