FitCF: A Framework for Automatic Feature Importance-guided Counterfactual Example Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Qianli, Feldhus, Nils, Ostermann, Simon, Villa-Arenas, Luis Felipe, Möller, Sebastian, Schmitt, Vera |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Cross-Refine: Improving Natural Language Explanation Generation by Learning in Tandem
by: Wang, Qianli, et al.
Published: (2024)
by: Wang, Qianli, et al.
Published: (2024)
Truth or Twist? Optimal Model Selection for Reliable Label Flipping Evaluation in LLM-based Counterfactuals
by: Wang, Qianli, et al.
Published: (2025)
by: Wang, Qianli, et al.
Published: (2025)
iFlip: Iterative Feedback-driven Counterfactual Example Refinement
by: Wang, Yilong, et al.
Published: (2026)
by: Wang, Yilong, et al.
Published: (2026)
CoXQL: A Dataset for Parsing Explanation Requests in Conversational XAI Systems
by: Wang, Qianli, et al.
Published: (2024)
by: Wang, Qianli, et al.
Published: (2024)
Can Large Language Models Still Explain Themselves? Investigating the Impact of Quantization on Self-Explanations
by: Wang, Qianli, et al.
Published: (2026)
by: Wang, Qianli, et al.
Published: (2026)
Parallel Universes, Parallel Languages: A Comprehensive Study on LLM-based Multilingual Counterfactual Example Generation
by: Wang, Qianli, et al.
Published: (2026)
by: Wang, Qianli, et al.
Published: (2026)
Through a Compressed Lens: Investigating The Impact of Quantization on Factual Knowledge Recall
by: Wang, Qianli, et al.
Published: (2025)
by: Wang, Qianli, et al.
Published: (2025)
Multilingual Datasets for Custom Input Extraction and Explanation Requests Parsing in Conversational XAI Systems
by: Wang, Qianli, et al.
Published: (2025)
by: Wang, Qianli, et al.
Published: (2025)
Anchored Alignment for Self-Explanations Enhancement
by: Villa-Arenas, Luis Felipe, et al.
Published: (2024)
by: Villa-Arenas, Luis Felipe, et al.
Published: (2024)
Judge Circuits
by: Feldhus, Nils, et al.
Published: (2026)
by: Feldhus, Nils, et al.
Published: (2026)
Proceedings of the ISCA/ITG Workshop on Diversity in Large Speech and Language Models
by: Möller, Sebastian, et al.
Published: (2025)
by: Möller, Sebastian, et al.
Published: (2025)
LLMCheckup: Conversational Examination of Large Language Models via Interpretability Tools and Self-Explanations
by: Wang, Qianli, et al.
Published: (2024)
by: Wang, Qianli, et al.
Published: (2024)
Enhancing Multilingual Counterfactual Generation through Alignment-as-Preference Optimization
by: Wang, Yilong, et al.
Published: (2026)
by: Wang, Yilong, et al.
Published: (2026)
What Are We Measuring in NLG? A Meta-Analysis of Evaluation Trends 2020-2025
by: Yang, Jing, et al.
Published: (2026)
by: Yang, Jing, et al.
Published: (2026)
Investigating the Interplay between Contextual and Parametric Chain-of-Thought Faithfulness under Optimization
by: Sun, Jingyi, et al.
Published: (2026)
by: Sun, Jingyi, et al.
Published: (2026)
Persona Prompting as a Lens on LLM Social Reasoning
by: Yang, Jing, et al.
Published: (2026)
by: Yang, Jing, et al.
Published: (2026)
Interpreting Language Models Through Concept Descriptions: A Survey
by: Feldhus, Nils, et al.
Published: (2025)
by: Feldhus, Nils, et al.
Published: (2025)
Free-text Rationale Generation under Readability Level Control
by: Hsu, Yi-Sheng, et al.
Published: (2024)
by: Hsu, Yi-Sheng, et al.
Published: (2024)
Simplifying Outcomes of Language Model Component Analyses with ELIA
by: Eidt, Aaron Louis, et al.
Published: (2026)
by: Eidt, Aaron Louis, et al.
Published: (2026)
Automatic Reviewers Fail to Detect Faulty Reasoning in Research Papers: A New Counterfactual Evaluation Framework
by: Dycke, Nils, et al.
Published: (2025)
by: Dycke, Nils, et al.
Published: (2025)
From Weights to Activations: Is Steering the Next Frontier of Adaptation?
by: Ostermann, Simon, et al.
Published: (2026)
by: Ostermann, Simon, et al.
Published: (2026)
Capturing Polysemanticity with PRISM: A Multi-Concept Feature Description Framework
by: Kopf, Laura, et al.
Published: (2025)
by: Kopf, Laura, et al.
Published: (2025)
Table Understanding and (Multimodal) LLMs: A Cross-Domain Case Study on Scientific vs. Non-Scientific Data
by: Borisova, Ekaterina, et al.
Published: (2025)
by: Borisova, Ekaterina, et al.
Published: (2025)
Gender Bias in Explainability: Investigating Performance Disparity in Post-hoc Methods
by: Dhaini, Mahdi, et al.
Published: (2025)
by: Dhaini, Mahdi, et al.
Published: (2025)
OpenFActScore: Open-Source Atomic Evaluation of Factuality in Text Generation
by: Lage, Lucas Fonseca, et al.
Published: (2025)
by: Lage, Lucas Fonseca, et al.
Published: (2025)
SCENE: Self-Labeled Counterfactuals for Extrapolating to Negative Examples
by: Fu, Deqing, et al.
Published: (2023)
by: Fu, Deqing, et al.
Published: (2023)
Hybrid Annotation for Propaganda Detection: Integrating LLM Pre-Annotations with Human Intelligence
by: Sahitaj, Ariana, et al.
Published: (2025)
by: Sahitaj, Ariana, et al.
Published: (2025)
Retrieving Climate Change Disinformation by Narrative
by: Upravitelev, Max, et al.
Published: (2026)
by: Upravitelev, Max, et al.
Published: (2026)
Why Does Reinforcement Learning Generalize? A Feature-Level Mechanistic Study of Post-Training in Large Language Models
by: Shi, Dan, et al.
Published: (2026)
by: Shi, Dan, et al.
Published: (2026)
Towards Automated Fact-Checking of Real-World Claims: Exploring Task Formulation and Assessment with LLMs
by: Sahitaj, Premtim, et al.
Published: (2025)
by: Sahitaj, Premtim, et al.
Published: (2025)
Generative Large Language Models in Automated Fact-Checking: A Survey
by: Vykopal, Ivan, et al.
Published: (2024)
by: Vykopal, Ivan, et al.
Published: (2024)
A Rigorous Evaluation of LLM Data Generation Strategies for Low-Resource Languages
by: Anikina, Tatiana, et al.
Published: (2025)
by: Anikina, Tatiana, et al.
Published: (2025)
Reverse Probing: Evaluating Knowledge Transfer via Finetuned Task Embeddings for Coreference Resolution
by: Anikina, Tatiana, et al.
Published: (2025)
by: Anikina, Tatiana, et al.
Published: (2025)
Infherno: End-to-end Agent-based FHIR Resource Synthesis from Free-form Clinical Notes
by: Frei, Johann, et al.
Published: (2025)
by: Frei, Johann, et al.
Published: (2025)
GrEmLIn: A Repository of Green Baseline Embeddings for 87 Low-Resource Languages Injected with Multilingual Graph Knowledge
by: Gurgurov, Daniil, et al.
Published: (2024)
by: Gurgurov, Daniil, et al.
Published: (2024)
The AI Language Proficiency Monitor -- Tracking the Progress of LLMs on Multilingual Benchmarks
by: Pomerenke, David, et al.
Published: (2025)
by: Pomerenke, David, et al.
Published: (2025)
Soft Language Prompts for Language Transfer
by: Vykopal, Ivan, et al.
Published: (2024)
by: Vykopal, Ivan, et al.
Published: (2024)
Engineering of Hallucination in Generative AI: It's not a Bug, it's a Feature
by: Fingscheidt, Tim, et al.
Published: (2026)
by: Fingscheidt, Tim, et al.
Published: (2026)
Does Using Counterfactual Help LLMs Explain Textual Importance in Classification?
by: Tan, Nelvin, et al.
Published: (2025)
by: Tan, Nelvin, et al.
Published: (2025)
AutoPsyC: Automatic Recognition of Psychodynamic Conflicts from Semi-structured Interviews with Large Language Models
by: Hossain, Sayed Muddashir, et al.
Published: (2025)
by: Hossain, Sayed Muddashir, et al.
Published: (2025)
Similar Items
-
Cross-Refine: Improving Natural Language Explanation Generation by Learning in Tandem
by: Wang, Qianli, et al.
Published: (2024) -
Truth or Twist? Optimal Model Selection for Reliable Label Flipping Evaluation in LLM-based Counterfactuals
by: Wang, Qianli, et al.
Published: (2025) -
iFlip: Iterative Feedback-driven Counterfactual Example Refinement
by: Wang, Yilong, et al.
Published: (2026) -
CoXQL: A Dataset for Parsing Explanation Requests in Conversational XAI Systems
by: Wang, Qianli, et al.
Published: (2024) -
Can Large Language Models Still Explain Themselves? Investigating the Impact of Quantization on Self-Explanations
by: Wang, Qianli, et al.
Published: (2026)