Towards Faithful Explanations: Boosting Rationalization with Shortcuts Discovery
Fuente:
arXiv
Saved in:
| Main Authors: | Yue, Linan, Liu, Qi, Du, Yichao, Wang, Li, Gao, Weibo, An, Yanqing |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Cooperative Classification and Rationalization for Graph Generalization
by: Yue, Linan, et al.
Published: (2024)
by: Yue, Linan, et al.
Published: (2024)
FaithLM: Towards Faithful Explanations for Large Language Models
by: Chuang, Yu-Neng, et al.
Published: (2024)
by: Chuang, Yu-Neng, et al.
Published: (2024)
Collaborative Cognitive Diagnosis with Disentangled Representation Learning for Learner Modeling
by: Gao, Weibo, et al.
Published: (2024)
by: Gao, Weibo, et al.
Published: (2024)
Event Grounded Criminal Court View Generation with Cooperative (Large) Language Models
by: Yue, Linan, et al.
Published: (2024)
by: Yue, Linan, et al.
Published: (2024)
DeepFaith: A Domain-Free and Model-Agnostic Unified Framework for Highly Faithful Explanations
by: Guo, Yuhan, et al.
Published: (2025)
by: Guo, Yuhan, et al.
Published: (2025)
Towards Few-shot Self-explaining Graph Neural Networks
by: Peng, Jingyu, et al.
Published: (2024)
by: Peng, Jingyu, et al.
Published: (2024)
Faithful and Stable Neuron Explanations for Trustworthy Mechanistic Interpretability
by: Yan, Ge, et al.
Published: (2025)
by: Yan, Ge, et al.
Published: (2025)
LIMEtree: Consistent and Faithful Surrogate Explanations of Multiple Classes
by: Sokol, Kacper, et al.
Published: (2020)
by: Sokol, Kacper, et al.
Published: (2020)
GraphPrompter: Multi-stage Adaptive Prompt Optimization for Graph In-Context Learning
by: Lv, Rui, et al.
Published: (2025)
by: Lv, Rui, et al.
Published: (2025)
QGShap: Quantum Acceleration for Faithful GNN Explanations
by: Jena, Haribandhu, et al.
Published: (2025)
by: Jena, Haribandhu, et al.
Published: (2025)
MiMu: Mitigating Multiple Shortcut Learning Behavior of Transformers
by: Zhao, Lili, et al.
Published: (2025)
by: Zhao, Lili, et al.
Published: (2025)
Training Multimodal Large Reasoning Models Needs Better Thoughts: A Three-Stage Framework for Long Chain-of-Thought Synthesis and Selection
by: Wang, Yizhi, et al.
Published: (2025)
by: Wang, Yizhi, et al.
Published: (2025)
A Positive Case for Faithfulness: LLM Self-Explanations Help Predict Model Behavior
by: Mayne, Harry, et al.
Published: (2026)
by: Mayne, Harry, et al.
Published: (2026)
Shortcut to Nowhere: Demystifying Deep Spurious Regression
by: Xu, Guanrong, et al.
Published: (2026)
by: Xu, Guanrong, et al.
Published: (2026)
Walk the Talk? Measuring the Faithfulness of Large Language Model Explanations
by: Matton, Katie, et al.
Published: (2025)
by: Matton, Katie, et al.
Published: (2025)
Learning with Logical Constraints but without Shortcut Satisfaction
by: Li, Zenan, et al.
Published: (2024)
by: Li, Zenan, et al.
Published: (2024)
Towards Metric-Faithful Neural Graph Matching
by: Shivottam, Jyotirmaya, et al.
Published: (2026)
by: Shivottam, Jyotirmaya, et al.
Published: (2026)
Towards Consistent Natural-Language Explanations via Explanation-Consistency Finetuning
by: Chen, Yanda, et al.
Published: (2024)
by: Chen, Yanda, et al.
Published: (2024)
Measuring the (Un)Faithfulness of Concept-Based Explanations
by: Kumar, Shubham, et al.
Published: (2025)
by: Kumar, Shubham, et al.
Published: (2025)
Inductive Subgraphs as Shortcuts: Causal Disentanglement for Heterophilic Graph Learning
by: Wang, Xiangmeng, et al.
Published: (2026)
by: Wang, Xiangmeng, et al.
Published: (2026)
Technical Note: Defining and Quantifying AND-OR Interactions for Faithful and Concise Explanation of DNNs
by: Li, Mingjie, et al.
Published: (2023)
by: Li, Mingjie, et al.
Published: (2023)
On the Complexity-Faithfulness Trade-off of Gradient-Based Explanations
by: Mehrpanah, Amir, et al.
Published: (2025)
by: Mehrpanah, Amir, et al.
Published: (2025)
Towards Interpretability Without Sacrifice: Faithful Dense Layer Decomposition with Mixture of Decoders
by: Oldfield, James, et al.
Published: (2025)
by: Oldfield, James, et al.
Published: (2025)
Towards Faithful Class-level Self-explainability in Graph Neural Networks by Subgraph Dependencies
by: Liu, Fanzhen, et al.
Published: (2025)
by: Liu, Fanzhen, et al.
Published: (2025)
Towards Verified and Targeted Explanations through Formal Methods
by: Wang, Hanchen David, et al.
Published: (2026)
by: Wang, Hanchen David, et al.
Published: (2026)
Gradient-based Model Shortcut Detection for Time Series Classification
by: Ibarra, Salomon, et al.
Published: (2025)
by: Ibarra, Salomon, et al.
Published: (2025)
Prior-Guided Symbolic Regression: Towards Scientific Consistency in Equation Discovery
by: Xiao, Jing, et al.
Published: (2026)
by: Xiao, Jing, et al.
Published: (2026)
Mitigating Shortcut Learning with InterpoLated Learning
by: Korakakis, Michalis, et al.
Published: (2025)
by: Korakakis, Michalis, et al.
Published: (2025)
Epistemic Traps: Rational Misalignment Driven by Model Misspecification
by: Xu, Xingcheng, et al.
Published: (2026)
by: Xu, Xingcheng, et al.
Published: (2026)
Towards Understanding the Influence of Training Samples on Explanations
by: Artelt, André, et al.
Published: (2024)
by: Artelt, André, et al.
Published: (2024)
Towards a Unified Framework for Evaluating Explanations
by: Pinto, Juan D., et al.
Published: (2024)
by: Pinto, Juan D., et al.
Published: (2024)
FaithfulSAE: Towards Capturing Faithful Features with Sparse Autoencoders without External Dataset Dependencies
by: Cho, Seonglae, et al.
Published: (2025)
by: Cho, Seonglae, et al.
Published: (2025)
Rational ANOVA Networks
by: Zhang, Jusheng, et al.
Published: (2026)
by: Zhang, Jusheng, et al.
Published: (2026)
Supervised Score-Based Modeling by Gradient Boosting
by: Zhao, Changyuan, et al.
Published: (2024)
by: Zhao, Changyuan, et al.
Published: (2024)
Rectifying Shortcut Behaviors in Preference-based Reward Learning
by: Ye, Wenqian, et al.
Published: (2025)
by: Ye, Wenqian, et al.
Published: (2025)
Towards Piece-by-Piece Explanations for Chess Positions with SHAP
by: Spinnato, Francesco
Published: (2025)
by: Spinnato, Francesco
Published: (2025)
Faithful Interpretation for Graph Neural Networks
by: Hu, Lijie, et al.
Published: (2024)
by: Hu, Lijie, et al.
Published: (2024)
Faithful or Just Plausible? Evaluating the Faithfulness of Closed-Source LLMs in Medical Reasoning
by: Afolabi, Halimat, et al.
Published: (2026)
by: Afolabi, Halimat, et al.
Published: (2026)
BEARS Make Neuro-Symbolic Models Aware of their Reasoning Shortcuts
by: Marconato, Emanuele, et al.
Published: (2024)
by: Marconato, Emanuele, et al.
Published: (2024)
Fairness is Not Flat: Geometric Phase Transitions Against Shortcut Learning
by: Rodriguez-Alvarez, Nicolas, et al.
Published: (2026)
by: Rodriguez-Alvarez, Nicolas, et al.
Published: (2026)
Similar Items
-
Cooperative Classification and Rationalization for Graph Generalization
by: Yue, Linan, et al.
Published: (2024) -
FaithLM: Towards Faithful Explanations for Large Language Models
by: Chuang, Yu-Neng, et al.
Published: (2024) -
Collaborative Cognitive Diagnosis with Disentangled Representation Learning for Learner Modeling
by: Gao, Weibo, et al.
Published: (2024) -
Event Grounded Criminal Court View Generation with Cooperative (Large) Language Models
by: Yue, Linan, et al.
Published: (2024) -
DeepFaith: A Domain-Free and Model-Agnostic Unified Framework for Highly Faithful Explanations
by: Guo, Yuhan, et al.
Published: (2025)