Probabilistic Stability Guarantees for Feature Attributions
Fuente:
arXiv
Saved in:
| Main Authors: | Jin, Helen, Xue, Anton, You, Weiqiu, Goel, Surbhi, Wong, Eric |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Sum-of-Parts: Self-Attributing Neural Networks with End-to-End Learning of Feature Groups
by: You, Weiqiu, et al.
Published: (2023)
by: You, Weiqiu, et al.
Published: (2023)
Probabilistic Soundness Guarantees in LLM Reasoning Chains
by: You, Weiqiu, et al.
Published: (2025)
by: You, Weiqiu, et al.
Published: (2025)
Logicbreaks: A Framework for Understanding Subversion of Rule-based Inference
by: Xue, Anton, et al.
Published: (2024)
by: Xue, Anton, et al.
Published: (2024)
The FIX Benchmark: Extracting Features Interpretable to eXperts
by: Jin, Helen, et al.
Published: (2024)
by: Jin, Helen, et al.
Published: (2024)
Reinforcement Learning for Control with Probabilistic Stability Guarantee: A Finite-Sample Approach
by: Han, Minghao, et al.
Published: (2026)
by: Han, Minghao, et al.
Published: (2026)
Why Do Transformers Fail to Forecast Time Series In-Context?
by: Zhou, Yufa, et al.
Published: (2025)
by: Zhou, Yufa, et al.
Published: (2025)
Model Agreement via Anchoring
by: Eaton, Eric, et al.
Published: (2026)
by: Eaton, Eric, et al.
Published: (2026)
Rigorous Probabilistic Guarantees for Robust Counterfactual Explanations
by: Marzari, Luca, et al.
Published: (2024)
by: Marzari, Luca, et al.
Published: (2024)
Missingness Bias Calibration in Feature Attribution Explanations
by: Sridhar, Shailesh, et al.
Published: (2026)
by: Sridhar, Shailesh, et al.
Published: (2026)
ANML: Attribution-Native Machine Learning with Guaranteed Robustness
by: Zahn, Oliver, et al.
Published: (2026)
by: Zahn, Oliver, et al.
Published: (2026)
Probabilistic Performance Guarantees for Multi-Task Reinforcement Learning
by: Schnitzer, Yannik, et al.
Published: (2026)
by: Schnitzer, Yannik, et al.
Published: (2026)
Counterfactual Explanations with Probabilistic Guarantees on their Robustness to Model Change
by: Stępka, Ignacy, et al.
Published: (2024)
by: Stępka, Ignacy, et al.
Published: (2024)
Probabilistic Dreaming for World Models
by: Wong, Gavin
Published: (2026)
by: Wong, Gavin
Published: (2026)
Towards Realistic Guarantees: A Probabilistic Certificate for SmoothLLM
by: Kumarappan, Adarsh, et al.
Published: (2025)
by: Kumarappan, Adarsh, et al.
Published: (2025)
Enumerating Safe Regions in Deep Neural Networks with Provable Probabilistic Guarantees
by: Marzari, Luca, et al.
Published: (2023)
by: Marzari, Luca, et al.
Published: (2023)
Impossibility Theorems for Feature Attribution
by: Bilodeau, Blair, et al.
Published: (2022)
by: Bilodeau, Blair, et al.
Published: (2022)
Feature Attribution Stability Suite: How Stable Are Post-Hoc Attributions?
by: Subramaniakuppusamy, Kamalasankari, et al.
Published: (2026)
by: Subramaniakuppusamy, Kamalasankari, et al.
Published: (2026)
Leveraging Approximate Model-based Shielding for Probabilistic Safety Guarantees in Continuous Environments
by: Goodall, Alexander W., et al.
Published: (2024)
by: Goodall, Alexander W., et al.
Published: (2024)
AXIL: Exact Instance Attribution for Gradient Boosting
by: Geertsema, Paul, et al.
Published: (2023)
by: Geertsema, Paul, et al.
Published: (2023)
On the Properties of Feature Attribution for Supervised Contrastive Learning
by: Arrighi, Leonardo, et al.
Published: (2026)
by: Arrighi, Leonardo, et al.
Published: (2026)
Correlation-Aware Feature Attribution Based Explainable AI
by: Sengupta, Poushali, et al.
Published: (2025)
by: Sengupta, Poushali, et al.
Published: (2025)
Hierarchical Delay Attribution Classification using Unstructured Text in Train Management Systems
by: Borg, Anton, et al.
Published: (2024)
by: Borg, Anton, et al.
Published: (2024)
A Theory of Learning with Autoregressive Chain of Thought
by: Joshi, Nirmit, et al.
Published: (2025)
by: Joshi, Nirmit, et al.
Published: (2025)
Distribution-Based Feature Attribution for Explaining the Predictions of Any Classifier
by: Li, Xinpeng, et al.
Published: (2025)
by: Li, Xinpeng, et al.
Published: (2025)
A Dual-Perspective Approach to Evaluating Feature Attribution Methods
by: Li, Yawei, et al.
Published: (2023)
by: Li, Yawei, et al.
Published: (2023)
A Polynomial-Time Axiomatic Alternative to SHAP for Feature Attribution
by: Hiraki, Kazuhiro, et al.
Published: (2026)
by: Hiraki, Kazuhiro, et al.
Published: (2026)
Conformal Alignment: Knowing When to Trust Foundation Models with Guarantees
by: Gui, Yu, et al.
Published: (2024)
by: Gui, Yu, et al.
Published: (2024)
Robust Counterfactual Explanations for Neural Networks With Probabilistic Guarantees
by: Hamman, Faisal, et al.
Published: (2023)
by: Hamman, Faisal, et al.
Published: (2023)
Adaptively profiling models with task elicitation
by: Brown, Davis, et al.
Published: (2025)
by: Brown, Davis, et al.
Published: (2025)
Generalized Attention Flow: Feature Attribution for Transformer Models via Maximum Flow
by: Azarkhalili, Behrooz, et al.
Published: (2025)
by: Azarkhalili, Behrooz, et al.
Published: (2025)
One Wave To Explain Them All: A Unifying Perspective On Feature Attribution
by: Kasmi, Gabriel, et al.
Published: (2024)
by: Kasmi, Gabriel, et al.
Published: (2024)
Hypothesis Class Determines Explanation: Why Accurate Models Disagree on Feature Attribution
by: B, Thackshanaramana
Published: (2026)
by: B, Thackshanaramana
Published: (2026)
On the Correlation between Individual Fairness and Predictive Accuracy in Probabilistic Models
by: Antonucci, Alessandro, et al.
Published: (2025)
by: Antonucci, Alessandro, et al.
Published: (2025)
A Pretrained Probabilistic Transformer for City-Scale Traffic Volume Prediction
by: Shen, Shiyu, et al.
Published: (2025)
by: Shen, Shiyu, et al.
Published: (2025)
Hybrid Attribution Priors for Explainable and Robust Model Training
by: Zhang, Zhuoran, et al.
Published: (2025)
by: Zhang, Zhuoran, et al.
Published: (2025)
DeepACTIF: Efficient Feature Attribution via Activation Traces in Neural Sequence Models
by: Hosp, Benedikt W.
Published: (2025)
by: Hosp, Benedikt W.
Published: (2025)
Causal SHAP: Feature Attribution with Dependency Awareness through Causal Discovery
by: Ng, Woon Yee, et al.
Published: (2025)
by: Ng, Woon Yee, et al.
Published: (2025)
MAC: A Conversion Rate Prediction Benchmark Featuring Labels Under Multiple Attribution Mechanisms
by: Wu, Jinqi, et al.
Published: (2026)
by: Wu, Jinqi, et al.
Published: (2026)
Stochastic Subgradient Methods with Guaranteed Global Stability in Nonsmooth Nonconvex Optimization
by: Xiao, Nachuan, et al.
Published: (2023)
by: Xiao, Nachuan, et al.
Published: (2023)
Shift-Invariant Feature Attribution in the Application of Wireless Electrocardiograms
by: Getnet, Yalemzerf, et al.
Published: (2026)
by: Getnet, Yalemzerf, et al.
Published: (2026)
Similar Items
-
Sum-of-Parts: Self-Attributing Neural Networks with End-to-End Learning of Feature Groups
by: You, Weiqiu, et al.
Published: (2023) -
Probabilistic Soundness Guarantees in LLM Reasoning Chains
by: You, Weiqiu, et al.
Published: (2025) -
Logicbreaks: A Framework for Understanding Subversion of Rule-based Inference
by: Xue, Anton, et al.
Published: (2024) -
The FIX Benchmark: Extracting Features Interpretable to eXperts
by: Jin, Helen, et al.
Published: (2024) -
Reinforcement Learning for Control with Probabilistic Stability Guarantee: A Finite-Sample Approach
by: Han, Minghao, et al.
Published: (2026)