Comparing Explanations is Not Enough, Explain the Change: New Standards are Needed to Explain Behavioral Shifts in Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Ciaperoni, Martino, Di Vece, Marzio, Pellungrini, Roberto, Pappalardo, Luca, Giannotti, Fosca, Giannini, Francesco |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Mathematical Foundation of Interpretable Equivariant Surrogate Models
by: Colombini, Jacopo Joy, et al.
Published: (2025)
by: Colombini, Jacopo Joy, et al.
Published: (2025)
AI, Meet Human: Learning Paradigms for Hybrid Decision Making Systems
by: Punzi, Clara, et al.
Published: (2024)
by: Punzi, Clara, et al.
Published: (2024)
Interpretable and Fair Mechanisms for Abstaining Classifiers
by: Lenders, Daphne, et al.
Published: (2025)
by: Lenders, Daphne, et al.
Published: (2025)
Interpretable and Fair Mechanisms for Abstaining Classifiers
by: Lenders, Daphne, et al.
Published: (2024)
by: Lenders, Daphne, et al.
Published: (2024)
Hybrid Retrieval for Hallucination Mitigation in Large Language Models: A Comparative Analysis
by: Mala, Chandana Sree, et al.
Published: (2025)
by: Mala, Chandana Sree, et al.
Published: (2025)
Deferring Concept Bottleneck Models: Learning to Defer Interventions to Inaccurate Experts
by: Pugnana, Andrea, et al.
Published: (2025)
by: Pugnana, Andrea, et al.
Published: (2025)
Logic Explanation of AI Classifiers by Categorical Explaining Functors
by: Fioravanti, Stefano, et al.
Published: (2025)
by: Fioravanti, Stefano, et al.
Published: (2025)
A Simulation Framework for Studying Systemic Effects of Feedback Loops in Recommender Systems
by: Barlacchi, Gabriele, et al.
Published: (2025)
by: Barlacchi, Gabriele, et al.
Published: (2025)
Reconciling econometrics with continuous maximum-entropy network models
by: Di Vece, Marzio, et al.
Published: (2022)
by: Di Vece, Marzio, et al.
Published: (2022)
Explanations Go Linear: Post-hoc Explainability for Tabular Data with Interpretable Meta-Encoding
by: Piaggesi, Simone, et al.
Published: (2025)
by: Piaggesi, Simone, et al.
Published: (2025)
Learning by Surprise: Surplexity for Mitigating Model Collapse in Generative AI
by: Gambetta, Daniele, et al.
Published: (2024)
by: Gambetta, Daniele, et al.
Published: (2024)
The Diversity Paradox revisited: Systemic Effects of Feedback Loops in Recommender Systems
by: Barlacchi, Gabriele, et al.
Published: (2026)
by: Barlacchi, Gabriele, et al.
Published: (2026)
Commodity-specific triads in the Dutch inter-industry production network
by: Di Vece, Marzio, et al.
Published: (2023)
by: Di Vece, Marzio, et al.
Published: (2023)
Explanations of Large Language Models Explain Language Representations in the Brain
by: Rahimi, Maryam, et al.
Published: (2025)
by: Rahimi, Maryam, et al.
Published: (2025)
A Survey on Graph Counterfactual Explanations: Definitions, Methods, Evaluation, and Research Challenges
by: Prado-Romero, Mario Alfonso, et al.
Published: (2022)
by: Prado-Romero, Mario Alfonso, et al.
Published: (2022)
Boosting Synthetic Data Generation with Effective Nonlinear Causal Discovery
by: Cinquini, Martina, et al.
Published: (2023)
by: Cinquini, Martina, et al.
Published: (2023)
Efficient Exploration of the Rashomon Set of Rule Set Models
by: Ciaperoni, Martino, et al.
Published: (2024)
by: Ciaperoni, Martino, et al.
Published: (2024)
Fair PCA, One Component at a Time
by: Matakos, Antonis, et al.
Published: (2025)
by: Matakos, Antonis, et al.
Published: (2025)
Sample and Expand: Discovering Low-rank Submatrices With Quality Guarantees
by: Ciaperoni, Martino, et al.
Published: (2025)
by: Ciaperoni, Martino, et al.
Published: (2025)
SMUG-Explain: A Framework for Symbolic Music Graph Explanations
by: Karystinaios, Emmanouil, et al.
Published: (2024)
by: Karystinaios, Emmanouil, et al.
Published: (2024)
Explainable Malware Detection with Tailored Logic Explained Networks
by: Anthony, Peter, et al.
Published: (2024)
by: Anthony, Peter, et al.
Published: (2024)
Explaining Explanation: An Empirical Study on Explanation in Code Reviews
by: Widyasari, Ratnadira, et al.
Published: (2023)
by: Widyasari, Ratnadira, et al.
Published: (2023)
Explaining Explanations in Probabilistic Logic Programming
by: Vidal, Germán
Published: (2024)
by: Vidal, Germán
Published: (2024)
Explaining GNN Explanations with Edge Gradients
by: He, Jesse, et al.
Published: (2025)
by: He, Jesse, et al.
Published: (2025)
Effective Explanations for Belief-Desire-Intention Robots: When and What to Explain
by: Wang, Cong, et al.
Published: (2025)
by: Wang, Cong, et al.
Published: (2025)
Perspectives in Play: A Multi-Perspective Approach for More Inclusive NLP Systems
by: Muscato, Benedetta, et al.
Published: (2025)
by: Muscato, Benedetta, et al.
Published: (2025)
Knowing What You Cannot Explain: Learning to Reject Low-Quality Explanations
by: Stradiotti, Luca, et al.
Published: (2025)
by: Stradiotti, Luca, et al.
Published: (2025)
LLMs for XAI: Future Directions for Explaining Explanations
by: Zytek, Alexandra, et al.
Published: (2024)
by: Zytek, Alexandra, et al.
Published: (2024)
Explaining Explaining
by: Nirenburg, Sergei, et al.
Published: (2024)
by: Nirenburg, Sergei, et al.
Published: (2024)
Explaining Autonomy: Enhancing Human-Robot Interaction through Explanation Generation with Large Language Models
by: Sobrín-Hidalgo, David, et al.
Published: (2024)
by: Sobrín-Hidalgo, David, et al.
Published: (2024)
Rate, Explain and Cite (REC): Enhanced Explanation and Attribution in Automatic Evaluation by Large Language Models
by: Hsu, Aliyah R., et al.
Published: (2024)
by: Hsu, Aliyah R., et al.
Published: (2024)
Can Large Language Models Still Explain Themselves? Investigating the Impact of Quantization on Self-Explanations
by: Wang, Qianli, et al.
Published: (2026)
by: Wang, Qianli, et al.
Published: (2026)
Khatri-Rao Clustering for Data Summarization
by: Ciaperoni, Martino, et al.
Published: (2026)
by: Ciaperoni, Martino, et al.
Published: (2026)
Counterintuitive Range Shifts May Be Explained by Climate Induced Changes in Biotic Interactions
by: Inna Osmolovsky, et al.
Published: (2025)
by: Inna Osmolovsky, et al.
Published: (2025)
Making a Change? Explain
Published: (2024)
Published: (2024)
Can Language Models Explain Their Own Classification Behavior?
by: Sherburn, Dane, et al.
Published: (2024)
by: Sherburn, Dane, et al.
Published: (2024)
Explaining Concept Shift with Interpretable Feature Attribution
by: Lyu, Ruiqi, et al.
Published: (2025)
by: Lyu, Ruiqi, et al.
Published: (2025)
GNN Explanations that do not Explain and How to find Them
by: Azzolin, Steve, et al.
Published: (2026)
by: Azzolin, Steve, et al.
Published: (2026)
Good Teachers Explain: Explanation-Enhanced Knowledge Distillation
by: Parchami-Araghi, Amin, et al.
Published: (2024)
by: Parchami-Araghi, Amin, et al.
Published: (2024)
Explaining k-Nearest Neighbors: Abductive and Counterfactual Explanations
by: Barceló, Pablo, et al.
Published: (2025)
by: Barceló, Pablo, et al.
Published: (2025)
Similar Items
-
Mathematical Foundation of Interpretable Equivariant Surrogate Models
by: Colombini, Jacopo Joy, et al.
Published: (2025) -
AI, Meet Human: Learning Paradigms for Hybrid Decision Making Systems
by: Punzi, Clara, et al.
Published: (2024) -
Interpretable and Fair Mechanisms for Abstaining Classifiers
by: Lenders, Daphne, et al.
Published: (2025) -
Interpretable and Fair Mechanisms for Abstaining Classifiers
by: Lenders, Daphne, et al.
Published: (2024) -
Hybrid Retrieval for Hallucination Mitigation in Large Language Models: A Comparative Analysis
by: Mala, Chandana Sree, et al.
Published: (2025)