DRIV-EX: Counterfactual Explanations for Driving LLMs
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Cardiel, Amaia, Zablocki, Eloi, Ramzi, Elias, Gaussier, Eric |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
GIFT: A Framework Towards Global Interpretable Faithful Textual Explanations of Vision Classifiers
par: Zablocki, Éloi, et autres
Publié: (2024)
par: Zablocki, Éloi, et autres
Publié: (2024)
LLM-wrapper: Black-Box Semantic-Aware Adaptation of Vision-Language Models for Referring Expression Comprehension
par: Cardiel, Amaia, et autres
Publié: (2024)
par: Cardiel, Amaia, et autres
Publié: (2024)
Paramanu: Compact and Competitive Monolingual Language Models for Low-Resource Morphologically Rich Indian Languages
par: Niyogi, Mitodru, et autres
Publié: (2024)
par: Niyogi, Mitodru, et autres
Publié: (2024)
Few-Shot Knowledge Distillation of LLMs With Counterfactual Explanations
par: Hamman, Faisal, et autres
Publié: (2025)
par: Hamman, Faisal, et autres
Publié: (2025)
Ayn: A Tiny yet Competitive Indian Legal Language Model Pretrained from Scratch
par: Niyogi, Mitodru, et autres
Publié: (2024)
par: Niyogi, Mitodru, et autres
Publié: (2024)
Guiding LLMs to Generate High-Fidelity and High-Quality Counterfactual Explanations for Text Classification
par: Nguyen, Van Bach, et autres
Publié: (2025)
par: Nguyen, Van Bach, et autres
Publié: (2025)
LIBERTy: A Causal Framework for Benchmarking Concept-Based Explanations of LLMs with Structural Counterfactuals
par: Toker, Gilat, et autres
Publié: (2026)
par: Toker, Gilat, et autres
Publié: (2026)
MAD: Motion Appearance Decoupling for efficient Driving World Models
par: Rahimi, Ahmad, et autres
Publié: (2026)
par: Rahimi, Ahmad, et autres
Publié: (2026)
Counterfactual Explanations for MITL Violations
par: Finkbeiner, Bernd, et autres
Publié: (2024)
par: Finkbeiner, Bernd, et autres
Publié: (2024)
Counterfactual Simulatability of LLM Explanations for Generation Tasks
par: Limpijankit, Marvin, et autres
Publié: (2025)
par: Limpijankit, Marvin, et autres
Publié: (2025)
LLMs Don't Know Their Own Decision Boundaries: The Unreliability of Self-Generated Counterfactual Explanations
par: Mayne, Harry, et autres
Publié: (2025)
par: Mayne, Harry, et autres
Publié: (2025)
Machine Unlearning Meets Adversarial Robustness via Constrained Interventions on LLMs
par: Rezkellah, Fatmazohra, et autres
Publié: (2025)
par: Rezkellah, Fatmazohra, et autres
Publié: (2025)
ReGentS: Real-World Safety-Critical Driving Scenario Generation Made Stable
par: Yin, Yuan, et autres
Publié: (2024)
par: Yin, Yuan, et autres
Publié: (2024)
A Comparative Analysis of Counterfactual Explanation Methods for Text Classifiers
par: McAleese, Stephen, et autres
Publié: (2024)
par: McAleese, Stephen, et autres
Publié: (2024)
Prompt-Counterfactual Explanations for Generative AI System Behavior
par: Goethals, Sofie, et autres
Publié: (2026)
par: Goethals, Sofie, et autres
Publié: (2026)
Attention Consistency for LLMs Explanation
par: Lan, Tian, et autres
Publié: (2025)
par: Lan, Tian, et autres
Publié: (2025)
Natural Language Counterfactual Explanations for Graphs Using Large Language Models
par: Giorgi, Flavio, et autres
Publié: (2024)
par: Giorgi, Flavio, et autres
Publié: (2024)
Explanation Generation for Contradiction Reconciliation with LLMs
par: Chan, Jason, et autres
Publié: (2026)
par: Chan, Jason, et autres
Publié: (2026)
Counterfactual Debating with Preset Stances for Hallucination Elimination of LLMs
par: Fang, Yi, et autres
Publié: (2024)
par: Fang, Yi, et autres
Publié: (2024)
Can LLMs Explain Themselves Counterfactually?
par: Dehghanighobadi, Zahra, et autres
Publié: (2025)
par: Dehghanighobadi, Zahra, et autres
Publié: (2025)
Aligning (Medical) LLMs for (Counterfactual) Fairness
par: Poulain, Raphael, et autres
Publié: (2024)
par: Poulain, Raphael, et autres
Publié: (2024)
ICR-Drive: Instruction Counterfactual Robustness for End-to-End Language-Driven Autonomous Driving
par: Hamid, Kaiser, et autres
Publié: (2026)
par: Hamid, Kaiser, et autres
Publié: (2026)
MEDEQUALQA: Evaluating Biases in LLMs with Counterfactual Reasoning
par: Ghosh, Rajarshi, et autres
Publié: (2025)
par: Ghosh, Rajarshi, et autres
Publié: (2025)
Investigating Counterfactual Unfairness in LLMs towards Identities through Humor
par: Kim, Shubin, et autres
Publié: (2026)
par: Kim, Shubin, et autres
Publié: (2026)
Argument-Based Consistency in Toxicity Explanations of LLMs
par: Mothilal, Ramaravind Kommiya, et autres
Publié: (2025)
par: Mothilal, Ramaravind Kommiya, et autres
Publié: (2025)
Do LLM Self-Explanations Help Users Predict Model Behavior? Evaluating Counterfactual Simulatability with Pragmatic Perturbations
par: Hong, Pingjun, et autres
Publié: (2026)
par: Hong, Pingjun, et autres
Publié: (2026)
Local Explanations and Self-Explanations for Assessing Faithfulness in black-box LLMs
par: Fragkathoulas, Christos, et autres
Publié: (2024)
par: Fragkathoulas, Christos, et autres
Publié: (2024)
LLMs for Generating and Evaluating Counterfactuals: A Comprehensive Study
par: Nguyen, Van Bach, et autres
Publié: (2024)
par: Nguyen, Van Bach, et autres
Publié: (2024)
Towards Unifying Evaluation of Counterfactual Explanations: Leveraging Large Language Models for Human-Centric Assessments
par: Domnich, Marharyta, et autres
Publié: (2024)
par: Domnich, Marharyta, et autres
Publié: (2024)
Success is in the Details: Evaluate and Enhance Details Sensitivity of Code LLMs through Counterfactuals
par: Luo, Xianzhen, et autres
Publié: (2025)
par: Luo, Xianzhen, et autres
Publié: (2025)
Generating Medically-Informed Explanations for Depression Detection using LLMs
par: Chen, Xiangyong, et autres
Publié: (2025)
par: Chen, Xiangyong, et autres
Publié: (2025)
Sensivity of LLMs' Explanations to the Training Randomness:Context, Class & Task Dependencies
par: Loncour, Romain, et autres
Publié: (2026)
par: Loncour, Romain, et autres
Publié: (2026)
Aligning What LLMs Do and Say: Towards Self-Consistent Explanations
par: Admoni, Sahar, et autres
Publié: (2025)
par: Admoni, Sahar, et autres
Publié: (2025)
Assessing the Reliability of LLMs Annotations in the Context of Demographic Bias and Model Explanation
par: Mohammadi, Hadi, et autres
Publié: (2025)
par: Mohammadi, Hadi, et autres
Publié: (2025)
Interpreting and Steering LLMs with Mutual Information-based Explanations on Sparse Autoencoders
par: Wu, Xuansheng, et autres
Publié: (2025)
par: Wu, Xuansheng, et autres
Publié: (2025)
Counterfactual Generation with Identifiability Guarantees
par: Yan, Hanqi, et autres
Publié: (2024)
par: Yan, Hanqi, et autres
Publié: (2024)
Counterfactual Cultural Cues Reduce Medical QA Accuracy in LLMs: Identifier vs Context Effects
par: Rezaei, Amirhossein Haji Mohammad, et autres
Publié: (2026)
par: Rezaei, Amirhossein Haji Mohammad, et autres
Publié: (2026)
Does Using Counterfactual Help LLMs Explain Textual Importance in Classification?
par: Tan, Nelvin, et autres
Publié: (2025)
par: Tan, Nelvin, et autres
Publié: (2025)
Digital Socrates: Evaluating LLMs through Explanation Critiques
par: Gu, Yuling, et autres
Publié: (2023)
par: Gu, Yuling, et autres
Publié: (2023)
Improving Implicit Discourse Relation Recognition with Natural Language Explanations from LLMs
par: Wang, Heng, et autres
Publié: (2026)
par: Wang, Heng, et autres
Publié: (2026)
Documents similaires
-
GIFT: A Framework Towards Global Interpretable Faithful Textual Explanations of Vision Classifiers
par: Zablocki, Éloi, et autres
Publié: (2024) -
LLM-wrapper: Black-Box Semantic-Aware Adaptation of Vision-Language Models for Referring Expression Comprehension
par: Cardiel, Amaia, et autres
Publié: (2024) -
Paramanu: Compact and Competitive Monolingual Language Models for Low-Resource Morphologically Rich Indian Languages
par: Niyogi, Mitodru, et autres
Publié: (2024) -
Few-Shot Knowledge Distillation of LLMs With Counterfactual Explanations
par: Hamman, Faisal, et autres
Publié: (2025) -
Ayn: A Tiny yet Competitive Indian Legal Language Model Pretrained from Scratch
par: Niyogi, Mitodru, et autres
Publié: (2024)