Impossibility Theorems for Feature Attribution
Fuente:
arXiv
Saved in:
| Main Authors: | Bilodeau, Blair, Jaques, Natasha, Koh, Pang Wei, Kim, Been |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Don't trust your eyes: on the (un)reliability of feature visualizations
by: Geirhos, Robert, et al.
Published: (2023)
by: Geirhos, Robert, et al.
Published: (2023)
Generative Modeling for Robust Deep Reinforcement Learning on the Traveling Salesman Problem
by: Li, Michael, et al.
Published: (2025)
by: Li, Michael, et al.
Published: (2025)
Maximizing Mutual Information Between Prompt and Response Improves LLM Performance With No Additional Data
by: Nam, Hyunji, et al.
Published: (2026)
by: Nam, Hyunji, et al.
Published: (2026)
Reliable, Adaptable, and Attributable Language Models with Retrieval
by: Asai, Akari, et al.
Published: (2024)
by: Asai, Akari, et al.
Published: (2024)
Modeling Others' Minds as Code
by: Jha, Kunal, et al.
Published: (2025)
by: Jha, Kunal, et al.
Published: (2025)
Learning to summarize user information for personalized reinforcement learning from human feedback
by: Nam, Hyunji, et al.
Published: (2025)
by: Nam, Hyunji, et al.
Published: (2025)
Infer Human's Intentions Before Following Natural Language Instructions
by: Wan, Yanming, et al.
Published: (2024)
by: Wan, Yanming, et al.
Published: (2024)
Learning to Cooperate with Humans using Generative Agents
by: Liang, Yancheng, et al.
Published: (2024)
by: Liang, Yancheng, et al.
Published: (2024)
On the Mathematical Impossibility of Safe Universal Approximators
by: Yao, Jasper
Published: (2025)
by: Yao, Jasper
Published: (2025)
Probabilistic Stability Guarantees for Feature Attributions
by: Jin, Helen, et al.
Published: (2025)
by: Jin, Helen, et al.
Published: (2025)
QuestBench: Can LLMs ask the right question to acquire information in reasoning tasks?
by: Li, Belinda Z., et al.
Published: (2025)
by: Li, Belinda Z., et al.
Published: (2025)
The Impossibility of Inverse Permutation Learning in Transformer Models
by: Alur, Rohan, et al.
Published: (2025)
by: Alur, Rohan, et al.
Published: (2025)
On the Properties of Feature Attribution for Supervised Contrastive Learning
by: Arrighi, Leonardo, et al.
Published: (2026)
by: Arrighi, Leonardo, et al.
Published: (2026)
Personalizing Reinforcement Learning from Human Feedback with Variational Preference Learning
by: Poddar, Sriyash, et al.
Published: (2024)
by: Poddar, Sriyash, et al.
Published: (2024)
Evaluating & Reducing Deceptive Dialogue From Language Models with Multi-turn RL
by: Abdulhai, Marwa, et al.
Published: (2025)
by: Abdulhai, Marwa, et al.
Published: (2025)
Correlation-Aware Feature Attribution Based Explainable AI
by: Sengupta, Poushali, et al.
Published: (2025)
by: Sengupta, Poushali, et al.
Published: (2025)
Mission: Impossible Language Models
by: Kallini, Julie, et al.
Published: (2024)
by: Kallini, Julie, et al.
Published: (2024)
Distribution-Based Feature Attribution for Explaining the Predictions of Any Classifier
by: Li, Xinpeng, et al.
Published: (2025)
by: Li, Xinpeng, et al.
Published: (2025)
A Dual-Perspective Approach to Evaluating Feature Attribution Methods
by: Li, Yawei, et al.
Published: (2023)
by: Li, Yawei, et al.
Published: (2023)
A Polynomial-Time Axiomatic Alternative to SHAP for Feature Attribution
by: Hiraki, Kazuhiro, et al.
Published: (2026)
by: Hiraki, Kazuhiro, et al.
Published: (2026)
Cross-environment Cooperation Enables Zero-shot Multi-agent Coordination
by: Jha, Kunal, et al.
Published: (2025)
by: Jha, Kunal, et al.
Published: (2025)
Evaluating the Predictive Features of Person-Centric Knowledge Graph Embeddings: Unfolding Ablation Studies
by: Theodoropoulos, Christos, et al.
Published: (2024)
by: Theodoropoulos, Christos, et al.
Published: (2024)
The Impossibility Triangle of Long-Context Modeling
by: Zhou, Yan
Published: (2026)
by: Zhou, Yan
Published: (2026)
The Features at Convergence Theorem: a first-principles alternative to the Neural Feature Ansatz for how networks learn representations
by: Boix-Adsera, Enric, et al.
Published: (2025)
by: Boix-Adsera, Enric, et al.
Published: (2025)
One Wave To Explain Them All: A Unifying Perspective On Feature Attribution
by: Kasmi, Gabriel, et al.
Published: (2024)
by: Kasmi, Gabriel, et al.
Published: (2024)
Generalized Attention Flow: Feature Attribution for Transformer Models via Maximum Flow
by: Azarkhalili, Behrooz, et al.
Published: (2025)
by: Azarkhalili, Behrooz, et al.
Published: (2025)
Hypothesis Class Determines Explanation: Why Accurate Models Disagree on Feature Attribution
by: B, Thackshanaramana
Published: (2026)
by: B, Thackshanaramana
Published: (2026)
Proximity-Informed Calibration for Deep Neural Networks
by: Xiong, Miao, et al.
Published: (2023)
by: Xiong, Miao, et al.
Published: (2023)
ALMo: Interactive Aim-Limit-Defined, Multi-Objective System for Personalized High-Dose-Rate Brachytherapy Treatment Planning and Visualization for Cervical Cancer
by: Chen, Edward, et al.
Published: (2026)
by: Chen, Edward, et al.
Published: (2026)
FreqX: Analyze the Attribution Methods in Another Domain
by: Liu, Zechen, et al.
Published: (2024)
by: Liu, Zechen, et al.
Published: (2024)
Sum-of-Parts: Self-Attributing Neural Networks with End-to-End Learning of Feature Groups
by: You, Weiqiu, et al.
Published: (2023)
by: You, Weiqiu, et al.
Published: (2023)
DeepACTIF: Efficient Feature Attribution via Activation Traces in Neural Sequence Models
by: Hosp, Benedikt W.
Published: (2025)
by: Hosp, Benedikt W.
Published: (2025)
VRAIL: Vectorized Reward-based Attribution for Interpretable Learning
by: Kim, Jina, et al.
Published: (2025)
by: Kim, Jina, et al.
Published: (2025)
Causal SHAP: Feature Attribution with Dependency Awareness through Causal Discovery
by: Ng, Woon Yee, et al.
Published: (2025)
by: Ng, Woon Yee, et al.
Published: (2025)
MAC: A Conversion Rate Prediction Benchmark Featuring Labels Under Multiple Attribution Mechanisms
by: Wu, Jinqi, et al.
Published: (2026)
by: Wu, Jinqi, et al.
Published: (2026)
Mission Impossible: A Statistical Perspective on Jailbreaking LLMs
by: Su, Jingtong, et al.
Published: (2024)
by: Su, Jingtong, et al.
Published: (2024)
Shift-Invariant Feature Attribution in the Application of Wireless Electrocardiograms
by: Getnet, Yalemzerf, et al.
Published: (2026)
by: Getnet, Yalemzerf, et al.
Published: (2026)
Modeling Multi-Objective Tradeoffs with Monotonic Utility Functions
by: Chen, Edward, et al.
Published: (2024)
by: Chen, Edward, et al.
Published: (2024)
Interaction-Aware Influence Functions for Group Attribution
by: Heo, Jaeseung, et al.
Published: (2026)
by: Heo, Jaeseung, et al.
Published: (2026)
Efficient Ensembles Improve Training Data Attribution
by: Deng, Junwei, et al.
Published: (2024)
by: Deng, Junwei, et al.
Published: (2024)
Similar Items
-
Don't trust your eyes: on the (un)reliability of feature visualizations
by: Geirhos, Robert, et al.
Published: (2023) -
Generative Modeling for Robust Deep Reinforcement Learning on the Traveling Salesman Problem
by: Li, Michael, et al.
Published: (2025) -
Maximizing Mutual Information Between Prompt and Response Improves LLM Performance With No Additional Data
by: Nam, Hyunji, et al.
Published: (2026) -
Reliable, Adaptable, and Attributable Language Models with Retrieval
by: Asai, Akari, et al.
Published: (2024) -
Modeling Others' Minds as Code
by: Jha, Kunal, et al.
Published: (2025)