ABLE: Using Adversarial Pairs to Construct Local Models for Explaining Model Predictions
Fuente:
arXiv
Saved in:
| Main Authors: | Khadka, Krishna, Shree, Sunny, Budhathoki, Pujan, Lei, Yu, Kacker, Raghu, Kuhn, D. Richard |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TabChange: Precise Attribute Changes in Tabular Data
by: Dahal, Arjun, et al.
Published: (2026)
by: Dahal, Arjun, et al.
Published: (2026)
DD-CAM: Minimal Sufficient Explanations for Vision Models Using Delta Debugging
by: Khadka, Krishna, et al.
Published: (2026)
by: Khadka, Krishna, et al.
Published: (2026)
TabKD: Tabular Knowledge Distillation through Interaction Diversity of Learned Feature Bins
by: Pereira, Shovon Niverd, et al.
Published: (2026)
by: Pereira, Shovon Niverd, et al.
Published: (2026)
LLM-Rank: A Graph Theoretical Approach to Pruning Large Language Models
by: Hoffmann, David, et al.
Published: (2024)
by: Hoffmann, David, et al.
Published: (2024)
Progressive Inference: Explaining Decoder-Only Sequence Classification Models Using Intermediate Predictions
by: Kariyappa, Sanjay, et al.
Published: (2024)
by: Kariyappa, Sanjay, et al.
Published: (2024)
FlavorDiffusion: Predicting Food Pairings and Chemical Interactions Using Diffusion Models
by: Pyo, Seo Jun
Published: (2025)
by: Pyo, Seo Jun
Published: (2025)
Physics-Informed Neural Network Surrogate Models for River Stage Prediction
by: Zoch, Maximilian, et al.
Published: (2025)
by: Zoch, Maximilian, et al.
Published: (2025)
In-Context Explainers: Harnessing LLMs for Explaining Black Box Models
by: Kroeger, Nicholas, et al.
Published: (2023)
by: Kroeger, Nicholas, et al.
Published: (2023)
Explaining Machine Learning Predictive Models through Conditional Expectation Methods
by: Ruiz-España, Silvia, et al.
Published: (2026)
by: Ruiz-España, Silvia, et al.
Published: (2026)
One-shot World Models Using a Transformer Trained on a Synthetic Prior
by: Ferreira, Fabio, et al.
Published: (2024)
by: Ferreira, Fabio, et al.
Published: (2024)
ShapShift: Explaining Model Prediction Shifts with Subgroup Conditional Shapley Values
by: Bewley, Tom, et al.
Published: (2026)
by: Bewley, Tom, et al.
Published: (2026)
Self-Explaining Hypergraph Neural Networks for Diagnosis Prediction
by: Yu, Leisheng, et al.
Published: (2025)
by: Yu, Leisheng, et al.
Published: (2025)
Explaining Predictions by Characteristic Rules
by: Alkhatib, Amr, et al.
Published: (2024)
by: Alkhatib, Amr, et al.
Published: (2024)
Explingo: Explaining AI Predictions using Large Language Models
by: Zytek, Alexandra, et al.
Published: (2024)
by: Zytek, Alexandra, et al.
Published: (2024)
Sequential Compression Layers for Efficient Federated Learning in Foundational Models
by: Mahla, Navyansh, et al.
Published: (2024)
by: Mahla, Navyansh, et al.
Published: (2024)
Helpful or Harmful Data? Fine-tuning-free Shapley Attribution for Explaining Language Model Predictions
by: Wang, Jingtan, et al.
Published: (2024)
by: Wang, Jingtan, et al.
Published: (2024)
Meaningful Causal Aggregation and Paradoxical Confounding
by: Zhu, Yuchen, et al.
Published: (2023)
by: Zhu, Yuchen, et al.
Published: (2023)
Explaining Large Language Models Decisions Using Shapley Values
by: Mohammadi, Behnam
Published: (2024)
by: Mohammadi, Behnam
Published: (2024)
A Model for Intelligible Interaction Between Agents That Predict and Explain
by: Baskar, A., et al.
Published: (2023)
by: Baskar, A., et al.
Published: (2023)
Adversarial Curriculum Graph Contrastive Learning with Pair-wise Augmentation
by: Zhao, Xinjian, et al.
Published: (2024)
by: Zhao, Xinjian, et al.
Published: (2024)
ConformaDecompose: Explaining Uncertainty via Calibration Localization
by: Yapicioglu, Fatima Rabia, et al.
Published: (2026)
by: Yapicioglu, Fatima Rabia, et al.
Published: (2026)
VeriDispatcher: Multi-Model Dispatching through Pre-Inference Difficulty Prediction for RTL Generation Optimization
by: Wang, Zeng, et al.
Published: (2025)
by: Wang, Zeng, et al.
Published: (2025)
Dreaming of Many Worlds: Learning Contextual World Models Aids Zero-Shot Generalization
by: Prasanna, Sai, et al.
Published: (2024)
by: Prasanna, Sai, et al.
Published: (2024)
Can Language Models Explain Their Own Classification Behavior?
by: Sherburn, Dane, et al.
Published: (2024)
by: Sherburn, Dane, et al.
Published: (2024)
Linearization Explains Fine-Tuning in Large Language Models
by: Afzal, Zahra Rahimi, et al.
Published: (2026)
by: Afzal, Zahra Rahimi, et al.
Published: (2026)
APT: Architectural Planning and Text-to-Blueprint Construction Using Large Language Models for Open-World Agents
by: Chen, Jun Yu, et al.
Published: (2024)
by: Chen, Jun Yu, et al.
Published: (2024)
Explaining Time Series via Contrastive and Locally Sparse Perturbations
by: Liu, Zichuan, et al.
Published: (2024)
by: Liu, Zichuan, et al.
Published: (2024)
Comparing Explanations is Not Enough, Explain the Change: New Standards are Needed to Explain Behavioral Shifts in Large Language Models
by: Ciaperoni, Martino, et al.
Published: (2026)
by: Ciaperoni, Martino, et al.
Published: (2026)
Inference Optimization of Foundation Models on AI Accelerators
by: Park, Youngsuk, et al.
Published: (2024)
by: Park, Youngsuk, et al.
Published: (2024)
Enhancing Adversarial Robustness with Conformal Prediction: A Framework for Guaranteed Model Reliability
by: Bao, Jie, et al.
Published: (2025)
by: Bao, Jie, et al.
Published: (2025)
Similarity-Based Self-Construct Graph Model for Predicting Patient Criticalness Using Graph Neural Networks and EHR Data
by: Sahu, Mukesh Kumar, et al.
Published: (2025)
by: Sahu, Mukesh Kumar, et al.
Published: (2025)
Explaining Predictive Uncertainty by Exposing Second-Order Effects
by: Bley, Florian, et al.
Published: (2024)
by: Bley, Florian, et al.
Published: (2024)
Use of What-if Scenarios to Help Explain Artificial Intelligence Models for Neonatal Health
by: Mamun, Abdullah, et al.
Published: (2024)
by: Mamun, Abdullah, et al.
Published: (2024)
TangledFeatures: Robust Feature Selection in Highly Correlated Spaces
by: Sunny, Allen Daniel
Published: (2025)
by: Sunny, Allen Daniel
Published: (2025)
Riemann Sum Optimization for Accurate Integrated Gradients Computation
by: Swain, Swadesh, et al.
Published: (2024)
by: Swain, Swadesh, et al.
Published: (2024)
MambaLRP: Explaining Selective State Space Sequence Models
by: Jafari, Farnoush Rezaei, et al.
Published: (2024)
by: Jafari, Farnoush Rezaei, et al.
Published: (2024)
Fairness Definitions in Language Models Explained
by: Yin, Zhipeng, et al.
Published: (2024)
by: Yin, Zhipeng, et al.
Published: (2024)
TopoReformer: Mitigating Adversarial Attacks Using Topological Purification in OCR Models
by: Kumar, Bhagyesh, et al.
Published: (2025)
by: Kumar, Bhagyesh, et al.
Published: (2025)
Using Mechanistic Interpretability to Craft Adversarial Attacks against Large Language Models
by: Winninger, Thomas, et al.
Published: (2025)
by: Winninger, Thomas, et al.
Published: (2025)
A Benchmark Time Series Dataset for Semiconductor Fabrication Manufacturing Constructed using Component-based Discrete-Event Simulation Models
by: Pendyala, Vamsi Krishna, et al.
Published: (2024)
by: Pendyala, Vamsi Krishna, et al.
Published: (2024)
Similar Items
-
TabChange: Precise Attribute Changes in Tabular Data
by: Dahal, Arjun, et al.
Published: (2026) -
DD-CAM: Minimal Sufficient Explanations for Vision Models Using Delta Debugging
by: Khadka, Krishna, et al.
Published: (2026) -
TabKD: Tabular Knowledge Distillation through Interaction Diversity of Learned Feature Bins
by: Pereira, Shovon Niverd, et al.
Published: (2026) -
LLM-Rank: A Graph Theoretical Approach to Pruning Large Language Models
by: Hoffmann, David, et al.
Published: (2024) -
Progressive Inference: Explaining Decoder-Only Sequence Classification Models Using Intermediate Predictions
by: Kariyappa, Sanjay, et al.
Published: (2024)